Principal Architect
This is us
Kaltura’s (NYSE:KLTR) mission is to power any video experience for any organization – live, on-demand, or real-time. We not only want to make using video simpler, but we also want to better people’s lives through video. Founded in 2006, Kaltura is now a global leader in the video market with millions of people using our products daily to teach, learn, watch, connect, and collaborate. Among our customers, you’ll find more than 1000 global, well-known organizations.
15+ years since starting the company, we continue to foster a diverse and collaborative work environment where everyone gets a say. Our team is currently 700+ people, and we’re still growing. We have offices in New York, London, Singapore, and Tel Aviv, but our technology is all in the cloud.
Kaltura has a fast-paced environment where initiative is always encouraged. Together with our hybrid work model and flexible state of mind, you get the right conditions for creative juices to flow freely. Thanks to our long line of products, cultivation of rich collaborative culture and care for each Kalturian, you’ll never run out of room to grow and evolve.
If you don't meet 100% of the requirements below - that's okay, nobody's perfect! We believe in hiring people, not just a list of skills. We encourage you to apply if you think this is a role that would make you excited about coming to work every day.
The Role
You own the end-to-end architecture of the agentic platform: how speech, avatar generation, and the agentic reasoning layer combine into one coherent system that holds up in production, across many tenants, under a real-time latency budget.
This is the widest technical seat in the group. A single user turn crosses speech recognition, retrieval, planning and tool calling, speech synthesis, and avatar rendering, and the experience is only as good as the seams between them. The scope here is the whole path, including the parts we did not build ourselves: externally sourced capabilities, third-party model providers, and customer-owned systems reached through the gateway.
Some of the decisions here shape the platform for a long time: how the speech pipeline is composed, which real-time orchestration and transport framework the platform standardizes on, how new capabilities are absorbed into the target architecture, and where a shared component belongs versus a per-customer one. You define the criteria these decisions are judged against and work closely with the research team that benchmarks the options, so the call rests on evidence. You write the reasoning down, and you stay close enough to the code to know when reality disagrees with the design.
You are hands-on. You prototype to de-risk, you read the traces yourself, and you stay as close to the numbers as to the diagram.
P
What You'll Do
Own the end-to-end system architecture across the real-time conversational path and the offline video path, defining the interfaces and boundaries between the speech, avatar, and agentic layers.
Drive the decisions that set platform direction, including speech pipeline composition, real-time orchestration and transport frameworks, model routing, and build-versus-adopt calls. You frame the criteria and the trade-offs, partner with the research team on the benchmarks that test them, and own the resulting call.
Lead the architecture for absorbing new capabilities, whether built, adopted, or acquired, and define the converged target architecture.
Architect for latency and cost as first-class constraints, owning the end-to-end latency budget, the cost-per-outcome model, and how the system degrades when a component stalls or fails.
Design the long-term memory architecture and the privacy model behind it: consent, retention, erasure, and residency.
Define cross-platform interfaces and engineering standards that let the foundation team and the FDE pods move independently, including the registry contract, hooks, and the governed path for customer-owned components.
Set the technical governance model for how capabilities are evaluated, promoted, versioned, and retired, with evaluation and tracing designed in rather than added afterward.
Prototype and de-risk future capabilities through focused technical spikes, turning results into clear recommendations.
Partner across the group and the company, from the research, platform, and FDE teams to product, and represent the architecture credibly to executives and customers.
r
What You Bring
Required
B.Sc. in Computer Science (or equivalent technical field), mandatory.
12+ years of industry experience, including 5+ years as a principal, staff, or lead architect owning system-level design for a production platform.
Proven track record architecting distributed, multi-tenant production systems at scale, with real accountability for latency, cost, and reliability.
Strong cloud experience, designing and running production systems on a major cloud platform.
Deep hands-on experience with agentic systems and LLMs, including orchestration, tool calling, interoperability standards such as MCP, retrieval, and how these systems fail in production.
Real-time systems experience: streaming transport, latency budgeting, and graceful degradation under load.
Strong Python skills. You still write code and read other people's code closely.
Hands-on experience with evaluation, tracing, and observability for AI systems, and using what they show to drive architectural change.
Strongly Preferred
Experience at a SaaS company.
Experience with speech technologies: ASR, TTS, or speech-to-speech, including streaming architectures.
Experience with generative video, avatar generation, or diffusion and flow-matching model families.
Experience with real-time media frameworks.
Experience with enterprise governance and compliance requirements.
Experience with multilingual systems.
M.Sc. or Ph.D. in Computer Science, Machine Learning, or a related field.
Required Skills
Required Languages
🇬🇧 English