About Inworld
Inworld is a research lab of top researchers and engineers, building the world’s top-ranked realtime voice models.
Today our models are the #1 ranked realtime voice models in the world. They are used to power the largest consumer-facing AI applications available, across categories like health, fitness, learning, therapy, companions, customer experience and media; representing 100s of millions of end users. Our work spans areas like research and development of state-of-the-art models, optimizing realtime inference, and creating best-in-class APIs and products that allow developers to engage their users.
We’ve raised more than $125M from Lightspeed, Section 32, Kleiner Perkins, Microsoft’s M12 venture fund, Founders Fund, Meta and Stanford, among others. Our technology has powered experiences from companies such as NVIDIA, Microsoft Xbox, Niantic, Logitech Streamlabs, Wishroll, Little Umbrella and Bible Chat. We’ve also been recognized by CB Insights as one of the 100 most promising AI companies globally and have been named one of LinkedIn’s Top 10 Startups in the USA.
About the role
Join our team as a Staff / Principal Platform Engineer and take end-to-end ownership of building, securing, and scaling our AI products. You'll be the driving force behind our cloud infrastructure, partnering with engineers across the organization to deploy and evolve services across major cloud providers using Terraform, ArgoCD, and other tooling. In this high-impact role, you'll identify what needs to be done and move it forward, directly shaping how we operate and innovate.
What you’ll do
Work closely with engineers to design, deploy, and maintain reliable, high-performance, and secure cloud infrastructure for our TTS and LLM Router.
Drive engineering velocity by identifying and building AI-powered tooling and workflows that improve how our teams develop and deploy software.
Facilitate a "you build it, you run it" culture by providing the necessary tools and processes for monitoring the reliability, availability, and performance of services.
Manage pipelines to ensure smooth and efficient code integration and deployment.
Conduct root cause analysis to identify critical issues and develop automated solutions to prevent recurrence.
Expected experience
8-10 years of experience in software engineering.
3+ years of experience with infrastructure-as-code.
Proficiency in managing Kubernetes clusters and applications, including creating Kustomize manifests/Helm charts for new applications.
Experience in creating and maintaining CI/CD pipelines for both applications and infrastructure deployments (using tools like Terraform/Terragrunt, ArgoCD, GitHub Actions, Ansible, etc.).
Deep knowledge of at least one major cloud provider (Google Cloud Platform, Microsoft Azure, Oracle Cloud).
Proficient in at least one backend programming/scripting languages such as Golang, Python, and Bash.
Candidates must be based in the SF Bay Area or willing to relocate (you will be working on-site in our South Bay office a few days a week).
The US base salary range for this full-time position is $280,000 - $350,000. In addition to base pay, total compensation includes equity and benefits. Within the range, individual pay is determined by work location, level, and additional factors, including competencies, experience, and business needs. The base pay range is subject to change and may be modified in the future.
Skills Required
- 8-10 years of software engineering experience
- 3+ years of experience with infrastructure-as-code
- Proficiency in managing Kubernetes clusters and applications
- Experience in creating and maintaining CI/CD pipelines
- Deep knowledge of at least one major cloud provider
- Proficient in at least one backend programming language
What We Do
Inworld AI is a realtime voice AI research lab and platform. It builds the voice layer for consumer AI, the speech that powers companions, tutors, coaches, customer-service agents, and creative and content applications. Inworld's flagship product is text-to-speech, top-ranked on the independent Artificial Analysis Speech Arena, and around it the company offers a full realtime voice stack. That stack includes text-to-speech (Realtime TTS-2, with over 100 languages and natural-language voice steering), speech-to-text, and an LLM router that reaches 200+ models from every major provider with no markup. Developers can call any product on its own, for example Realtime TTS through a single API call, or combine speech-to-text, an LLM, and text-to-speech into one live conversation through the Realtime API. Dedicated GPUs are available for the largest workloads. Inworld serves developers and founders building consumer AI products. As voice quality becomes commoditized, the company's focus is realtime conversation at scale: voice agents that respond in roughly 200 milliseconds for thousands of simultaneous users, at a cost that stays efficient as usage grows. Inworld is a product-oriented research lab. Its founding team pioneered conversational AI and generative models at API.AI (acquired by Google and renamed Dialogflow), Google, and DeepMind, with expertise spanning language models, speech synthesis, multimodal interaction, and design. Inworld has raised more than $125M from investors including Lightspeed, Kleiner Perkins, Founders Fund, CRV, Intel Capital, BITKRAFT Ventures, Section 32, Meta, Microsoft's M12, and LG Technology Ventures. The company was one of six selected for the 2022 Disney Accelerator.

.png)







