AI Engineer

Posted Yesterday
Be an Early Applicant
Hiring Remotely in Israel
Remote or Hybrid
Mid level
Cloud • Edtech • Marketing Tech • Software
Kaltura provides both live and on-demand video solutions that bring brands and people together.
The Role
Build the shared platform layer for Kaltura’s agentic AI systems, including orchestration, skills registries, evaluation, tracing, guardrails, memory, retrieval, and integration gateways. Develop production backend services in Python supporting real-time and offline video runtimes. Design scalable, multi-tenant platform primitives, integrate LLM and RAG infrastructure, optimize for latency, and improve reliability, observability, safety, and customer extensibility.
Summary Generated by Built In
Description

This is us 

Kaltura’s (NYSE:KLTR) mission is to power any video experience for any organization – live, on-demand, or real-time. We not only want to make using video simpler, but we also want to better people’s lives through video. Founded in 2006, Kaltura is now a global leader in the video market with millions of people using our products daily to teach, learn, watch, connect, and collaborate. Among our customers, you’ll find more than 1000 global, well-known organizations.    

15+ years since starting the company, we continue to foster a diverse and collaborative work environment where everyone gets a say. Our team is currently 700+ people, and we’re still growing. We have offices in New York, London, Singapore, and Tel Aviv, but our technology is all in the cloud. 

Kaltura has a fast-paced environment where initiative is always encouraged. Together with our hybrid work model and flexible state of mind, you get the right conditions for creative juices to flow freely. Thanks to our long line of products, cultivation of rich collaborative culture and care for each Kalturian, you’ll never run out of room to grow and evolve.   

If you don't meet 100% of the requirements below - that's okay, nobody's perfect! We believe in hiring people, not just a list of skills. We encourage you to apply if you think this is a role that would make you excited about coming to work every day. 

Requirements

The Role

You will build the shared platform layer that makes Kaltura's agentic AI extensible, measurable, and self-improving. This is the foundation beneath everything: the orchestration engine, the skills registry, the evaluation harness, the guardrails system, the memory service, the integration gateway, and the tracing infrastructure — consumed by both the real-time conversational and offline video generation runtimes.

Your work is what separates "it works for one customer with engineering involvement" from "it works for many customers across many domains without bespoke engineering per customer." You build the machinery once; customer-facing teams and eventually customers themselves extend it.

Scope

You will work across the platform's core systems: agentic orchestration (planning, sub-agents, parallel execution under deadlines), skills registry, evaluation and tracing, guardrails, integration gateway, memory service, and retrieval infrastructure. The shared machinery that both the real-time and offline runtimes consume.

What You Bring

Required

· 4+ years building production backend/platform systems — distributed services, APIs, async processing, and systems that serve multiple consumers. Python primary; experience with high-throughput, low-latency services.

· Deep hands-on experience with LLMs in production — not just using them, but building the infrastructure around them: orchestration, tool calling, retrieval pipelines, context management, prompt chaining, and multi-agent coordination. Proficiency with LangChain, LangGraph, or equivalent orchestration frameworks.

· Experience building platform primitives — registries, gateways, evaluation frameworks, tracing systems, or similar shared infrastructure that other teams build on top of. You understand contracts, versioning, and what it means to ship a platform rather than a feature.

· Demonstrated ability to make build vs. adopt decisions — you have integrated open-source tooling into production systems, understood its boundaries, and built the proprietary layer where needed.

· Experience with RAG systems at depth — indexing strategies, retrieval evaluation, chunking, re-ranking, hybrid search, and the failure modes (context pollution, retrieval misses, conflicting sources). Familiarity with vector databases (Pinecone, Weaviate, Qdrant, or similar) and embedding models.

· Familiarity with real-time system constraints — you understand how latency budgets, deadlines, and streaming shape architecture differently than batch systems.

· Strong systems design skills — you can design a registry, a gateway, a memory service, or an evaluation harness that serves multiple consumers with different needs (internal teams and external customers, real-time and offline paths).

Strongly Preferred

· Experience with evaluation of LLM systems — designing metrics, building test harnesses, measuring behavioral consistency, detecting regressions. Familiarity with DeepEval, Ragas, or similar frameworks.

· Experience with multi-tenant platforms — tenant isolation, scoped configuration, and the engineering discipline required when one customer's data/behavior must never leak to another.

· Background in speech/multimodal AI pipelines — ASR, TTS, avatar rendering, or video generation pipelines. Understanding how these compose with a reasoning layer.

· Experience with observability and tracing at the platform level — structured per-decision traces, not just application logs. Familiarity with Langfuse, Opik, OpenTelemetry, or similar.

· Familiarity with agentic AI frameworks and patterns — CrewAI, AutoGen, Semantic Kernel, MCP protocol, and the failure modes unique to multi-step AI systems (planning loops, tool selection, delegation, parallel execution).

· Experience with guardrails and safety for AI systems — input filtering, output validation, grounding checks, and the latency trade-offs of validation in real-time streaming.

What Success Looks Like

1 month: Deep understanding of the current architecture (Brain, Conversation Manager, real-time and offline pipelines). First contribution shipped — either to evaluation/tracing adoption or registry infrastructure.

3 months: One major platform component owned and in production — registry, evaluation harness, guardrails, or gateway. Integrated with the real-time pipeline and validated under latency constraints. FDE teams consuming your work.

6 months: Multiple foundation components live and serving both runtimes. Evaluation gates enforced on registry entries. Tracing producing actionable diagnostics. Your components are what make new customer onboardings faster rather than bespoke.

Why This Role

You will build the platform layer that turns proven AI technology into an enterprise-grade system. The engine exists — real-time avatar conversations, low-latency speech, offline video generation. What doesn't exist yet is the shared infrastructure that makes it extensible, measurable, and reliable across many customers. That's what you build.

Every component you ship is consumed by multiple teams and multiple customers. Your work compounds.

Skills Required

  • 4+ years building production backend or platform systems, including distributed services, APIs, asynchronous processing, and multi-consumer systems
  • Strong Python experience and experience building high-throughput, low-latency services
  • Hands-on production experience with LLM infrastructure, including orchestration, tool calling, retrieval pipelines, context management, prompt chaining, and multi-agent coordination
  • Proficiency with LangChain, LangGraph, or equivalent orchestration frameworks
  • Experience building platform primitives such as registries, gateways, evaluation frameworks, or tracing systems
  • Experience making build-versus-adopt decisions and integrating open-source tooling into production systems
  • Deep experience with RAG systems, including indexing, retrieval evaluation, chunking, reranking, hybrid search, and retrieval failure modes
  • Familiarity with vector databases such as Pinecone, Weaviate, Qdrant, or similar, and embedding models
  • Understanding of real-time system constraints, latency budgets, deadlines, and streaming architectures
  • Strong systems design skills for shared services serving internal teams and external customers
  • Experience evaluating LLM systems, designing metrics, building test harnesses, measuring consistency, and detecting regressions
  • Familiarity with DeepEval, Ragas, or similar evaluation frameworks
  • Experience with multi-tenant platforms, tenant isolation, and scoped configuration
  • Background in speech or multimodal AI pipelines, including ASR, TTS, avatar rendering, or video generation
  • Experience with platform-level observability and tracing
  • Familiarity with Langfuse, Opik, OpenTelemetry, or similar observability tools
  • Familiarity with agentic AI frameworks and patterns such as CrewAI, AutoGen, Semantic Kernel, or MCP
  • Experience with AI guardrails and safety, including input filtering, output validation, and grounding checks
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: New York, NY
730 Employees
Year Founded: 2006

What We Do

Kaltura’s mission is to power any video experience for any organization. Kaltura is the leading video cloud, powering the broadest range of video experiences. Kaltura’s products are used by thousands of global enterprises, media companies, service providers and educational institutions, engaging hundreds of millions of viewers at home, at work, and at school. More than 15 years after starting this company, we continue to foster a diverse and collaborative work environment where everyone gets a say. Together with our hybrid work model and flexible state of mind, you get the right conditions for creative juices to flow freely. What's more, even with almost 1000 professionals around the world, our DNA has remained intact. Each new person who joins this company is just like the original Kalturians: smart, empathetic, daring, and intrapreneurial at heart.

Why Work With Us

Thanks to our long line of products, cultivation of rich collaborative culture and care for each Kalturian, you’ll never run out of room to grow and evolve. In fact, we prize internal mobility and personal growth. No matter what, we’ll encourage you to follow your gut feelings, learn new things, and take full ownership of your ideas.

Gallery

Gallery

Similar Jobs

HiBob Logo HiBob

Senior AI Ops Engineer

HR Tech • Information Technology • Professional Services • Sales • Software
Remote or Hybrid
IL
1350 Employees

Micron Technology Logo Micron Technology

Artificial Intelligence Engineer

Artificial Intelligence • Hardware • Information Technology • Machine Learning
In-Office or Remote
Hadera, ISR
45000 Employees

HiBob Logo HiBob

Senior Back-end Engineer

HR Tech • Information Technology • Professional Services • Sales • Software
Remote or Hybrid
IL
1350 Employees

Mobileye Logo Mobileye

Platform Engineer

Automotive • Automation
Remote or Hybrid
Jerusalem, ISR
3700 Employees

Similar Companies Hiring

Kepler  Thumbnail
Artificial Intelligence • Fintech • Software
New York, New York
9 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees
Revel.io Thumbnail
Aerospace • Hardware • Robotics • Software
US
50 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account