This is us
Kaltura’s (NYSE:KLTR) mission is to power any video experience for any organization – live, on-demand, or real-time. We not only want to make using video simpler, but we also want to better people’s lives through video. Founded in 2006, Kaltura is now a global leader in the video market with millions of people using our products daily to teach, learn, watch, connect, and collaborate. Among our customers, you’ll find more than 1000 global, well-known organizations.
15+ years since starting the company, we continue to foster a diverse and collaborative work environment where everyone gets a say. Our team is currently 700+ people, and we’re still growing. We have offices in New York, London, Singapore, and Tel Aviv, but our technology is all in the cloud.
Kaltura has a fast-paced environment where initiative is always encouraged. Together with our hybrid work model and flexible state of mind, you get the right conditions for creative juices to flow freely. Thanks to our long line of products, cultivation of rich collaborative culture and care for each Kalturian, you’ll never run out of room to grow and evolve.
If you don't meet 100% of the requirements below - that's okay, nobody's perfect! We believe in hiring people, not just a list of skills. We encourage you to apply if you think this is a role that would make you excited about coming to work every day.
What you'll do
You will lead the DevOps Infrastructure & Developer Experience team — the team that owns the foundational layer every engineering team at Kaltura builds on. EKS clusters, VPCs, IAM, security boundaries, monitoring, CI/CD, and the developer workflows that tie it all together. When your infrastructure works well, every team ships faster. When it doesn't, everyone feels it.
You'll manage a team of DevOps engineers, set technical direction, and own the infrastructure platform end-to-end — from networking and compute through security and observability to the developer experience layer on top. You are accountable for the reliability, security, and usability of the shared platform, and you measure success by the outcomes it enables across the organization.
This role sits at the intersection of infrastructure depth and organizational breadth. You'll partner with Development, Platform, AI/ML, Security, and Product — ensuring the underlying systems evolve to meet their needs while maintaining the standards and guardrails that keep production safe.
We're looking for someone who has already done this — operated production infrastructure at scale, shipped developer tooling, driven AI adoption in engineering workflows, and measured the impact of
The day-to-day
Team Leadership
Lead and grow a team of DevOps engineers - set direction, remove blockers, own the roadmap
Balance operational needs with platform investment. Represent infrastructure tradeoffs in cross-org planning
Drive hiring, onboarding, and professional growth
Infrastructure Ownership
Own the shared layer all teams depend on - EKS, VPC, IAM, networking, security, compute (including GPU), monitoring, data services
Design and operate highly available distributed systems at scale on AWS and GCP. IaC with Terraform, multi-account, multi-region
Own security posture (IAM, network boundaries, secrets, vulnerability scanning) and observability (Prometheus, Grafana, CloudWatch, alerting)
Developer Experience & Shipping
Own how developers interact with infrastructure - the platform is your product, developers are your users
Own the full path from commit to production - CI/CD (GitHub Actions), environment promotion, progressive rollouts, automated rollback
Build golden paths and self-service workflows. If a developer has to ask your team twice, automate it
Make shipping fast and safe by default — security scanning, policy enforcement, and blast radius controls baked into the pipeline
AI Agents & AI-Native Development
Build and operate infrastructure for AI workloads - GPU clusters, inference serving, model deployment
Enable AI agent adoption across engineering - execution environments, tooling, guardrails
Champion AI coding tools within the team and across the org. Apply agents to infrastructure operations
Measuring Outcomes
Define success metrics for every initiative. Measure before and after
Own platform KPIs: reliability, developer velocity (deploy frequency, lead time, PR cycle time), cost efficiency, security posture
Use data to prioritize — invest where developers are slowest or the platform is weakest
Ideally, we’re looking for:
4+ years in DevOps / Platform / SRE roles, managing a team, operating in a SaaS production environment with multi-region deployment within an R&D organization
Deep hands-on expertise with AWS in production — EKS, VPC, IAM, EC2, S3, CloudFront, MSK, Lambda, CloudWatch, Security Hub. You've built and owned multi-account infrastructure, not just used it
Experience with GCP in production — GKE, Cloud Run, IAM, VPC, Cloud Build, or similar. Comfortable operating across cloud providers
Deep expertise with Kubernetes — cluster operations, networking (CNI, service mesh), RBAC, node lifecycle, Helm chart architecture. You understand the internals, not just the YAML
Strong networking and OS fundamentals — TCP/IP, DNS, load balancing, Linux internals, troubleshooting at every layer
Proven experience owning developer experience — you've built internal platforms, CI/CD systems, or self-service tooling that developers actually adopted and that measurably improved their velocity
Strong security mindset — IAM design, network segmentation. Security is part of how you build, not an afterthought
Experience with CI/CD at scale (GitHub Actions), monitoring and observability (Prometheus, Grafana), and Infrastructure as Code (Terraform)
High proficiency with AI coding tools - you use them daily, understand how they work, and can drive their adoption across a team
These would also be nice:
Experience building or operating GPU infrastructure for AI/ML inference at scale
Experience with AI agents — building them, deploying them in production, or enabling teams to use them
Experience consolidating or migrating infrastructure across acquisitions or organizational changes
Experience defining and reporting on engineering metrics (DORA, platform KPIs, cost models)
Experience with multi-cloud (AWS + GCP) infrastructure
The perks:
Hybrid, flexible work environment
Extended private health (including mental) insurance