Kaltura

DevOps Team Leader - Infra and DevEx

Kaltura  •  Bnei Brak, IL (Hybrid)  •  22 hours ago
Apply
AI can make mistakes so check important info. Chat history is never stored.

Job Description

This is us

Kaltura’s (NYSE:KLTR) mission is to power any video experience for any organization – live, on-demand, or real-time. We not only want to make using video simpler, but we also want to better people’s lives through video. Founded in 2006, Kaltura is now a global leader in the video market with millions of people using our products daily to teach, learn, watch, connect, and collaborate. Among our customers, you’ll find more than 1000 global, well-known organizations.

15+ years since starting the company, we continue to foster a diverse and collaborative work environment where everyone gets a say. Our team is currently 700+ people, and we’re still growing. We have offices in New York, London, Singapore, and Tel Aviv, but our technology is all in the cloud.

Kaltura has a fast-paced environment where initiative is always encouraged. Together with our hybrid work model and flexible state of mind, you get the right conditions for creative juices to flow freely. Thanks to our long line of products, cultivation of rich collaborative culture and care for each Kalturian, you’ll never run out of room to grow and evolve.

If you don't meet 100% of the requirements below - that's okay, nobody's perfect! We believe in hiring people, not just a list of skills. We encourage you to apply if you think this is a role that would make you excited about coming to work every day.

Requirements

The role

You will lead the DevOps Infrastructure & Developer Experience team — the team that owns the foundational layer every engineering team at Kaltura builds on. EKS clusters, VPCs, IAM, security boundaries, monitoring, CI/CD, and the developer workflows that tie it all together. When your infrastructure works well, every team ships faster. When it doesn't, everyone feels it.

You'll manage a team of DevOps engineers, set technical direction, and own the infrastructure platform end-to-end — from networking and compute through security and observability to the developer experience layer on top. You are accountable for the reliability, security, and usability of the shared platform, and you measure success by the outcomes it enables across the organization.

This role sits at the intersection of infrastructure depth and organizational breadth. You'll partner with Development, Platform, AI/ML, Security, and Product — ensuring the underlying systems evolve to meet their needs while maintaining the standards and guardrails that keep production safe.

We're looking for someone who has already done this — operated production infrastructure at scale, shipped developer tooling, driven AI adoption in engineering workflows, and measured the impact of

The day-to-day

Team Leadership

  • Lead and grow a team of DevOps engineers - set direction, remove blockers, own the roadmap
  • Balance operational needs with platform investment. Represent infrastructure tradeoffs in cross-org planning
  • Drive hiring, onboarding, and professional growth

Infrastructure Ownership

  • Own the shared layer all teams depend on - EKS, VPC, IAM, networking, security, compute (including GPU), monitoring, data services
  • Design and operate highly available distributed systems at scale on AWS and GCP. IaC with Terraform, multi-account, multi-region
  • Own security posture (IAM, network boundaries, secrets, vulnerability scanning) and observability (Prometheus, Grafana, CloudWatch, alerting)

Developer Experience & Shipping

  • Own how developers interact with infrastructure - the platform is your product, developers are your users
  • Own the full path from commit to production - CI/CD (GitHub Actions), environment promotion, progressive rollouts, automated rollback
  • Build golden paths and self-service workflows. If a developer has to ask your team twice, automate it
  • Make shipping fast and safe by default — security scanning, policy enforcement, and blast radius controls baked into the pipeline

AI Agents & AI-Native Development

  • Build and operate infrastructure for AI workloads - GPU clusters, inference serving, model deployment
  • Enable AI agent adoption across engineering - execution environments, tooling, guardrails
  • Champion AI coding tools within the team and across the org. Apply agents to infrastructure operations

Measuring Outcomes

  • Define success metrics for every initiative. Measure before and after
  • Own platform KPIs: reliability, developer velocity (deploy frequency, lead time, PR cycle time), cost efficiency, security posture
  • Use data to prioritize — invest where developers are slowest or the platform is weakest

Ideally, we’re looking for:

  • 4+ years in DevOps / Platform / SRE roles, managing a team, operating in a SaaS production environment with multi-region deployment within an R&D organization
  • Deep hands-on expertise with AWS in production — EKS, VPC, IAM, EC2, S3, CloudFront, MSK, Lambda, CloudWatch, Security Hub. You've built and owned multi-account infrastructure, not just used it
  • Experience with GCP in production — GKE, Cloud Run, IAM, VPC, Cloud Build, or similar. Comfortable operating across cloud providers
  • Deep expertise with Kubernetes — cluster operations, networking (CNI, service mesh), RBAC, node lifecycle, Helm chart architecture. You understand the internals, not just the YAML
  • Strong networking and OS fundamentals — TCP/IP, DNS, load balancing, Linux internals, troubleshooting at every layer
  • Proven experience owning developer experience — you've built internal platforms, CI/CD systems, or self-service tooling that developers actually adopted and that measurably improved their velocity
  • Strong security mindset — IAM design, network segmentation. Security is part of how you build, not an afterthought
  • Experience with CI/CD at scale (GitHub Actions), monitoring and observability (Prometheus, Grafana), and Infrastructure as Code (Terraform)
  • High proficiency with AI coding tools - you use them daily, understand how they work, and can drive their adoption across a team

These would also be nice:

  • Experience building or operating GPU infrastructure for AI/ML inference at scale
  • Experience with AI agents — building them, deploying them in production, or enabling teams to use them
  • Experience consolidating or migrating infrastructure across acquisitions or organizational changes
  • Experience defining and reporting on engineering metrics (DORA, platform KPIs, cost models)
  • Experience with multi-cloud (AWS + GCP) infrastructure

The perks:

  1. Hybrid, flexible work environment
  2. Extended private health (including mental) insurance
  3. Personal and professional development programs
  4. Occasional Cross company long weekends
Kaltura

About Kaltura

Kaltura’s mission is to power any video experience for any organization. Kaltura is the leading video cloud, powering the broadest range of video experiences. Our Video Experience Cloud is used by leading brands reaching millions of users, at home, at school, and at work, for virtual events, communication, collaboration, training, marketing, sales, customer care, teaching, learning, and entertainment experiences.

Industry
IT & Software
Company Size
501-1,000 employees
Headquarters
Remote First
Year Founded
2006
Social Media