Staff / Principal Platform Engineer - USA

Reposted One Month Ago
Be an Early Applicant
Mountain View, CA, USA
Hybrid
280K-350K Annually
Senior level
Software
The Role
As a Staff/Principal Platform Engineer, you will build, secure, and scale AI products, manage cloud infrastructure, integrate CI/CD pipelines, and enhance developer workflows with automated tooling.
Summary Generated by Built In

About Inworld

Inworld is a research lab and inference provider focused on realtime AI for consumer-facing applications. We build first-party speech models, serve LLMs, and run the inference behind modular APIs designed for high-volume, realtime workloads.

Hundreds of millions of users interact with Inworld powered apps every day and we serve over 10 trillion LLM tokens per month. Our models and infrastructure support consumer applications across companions, healthcare, fitness, education, media, and more. Our work spans model research, realtime inference, large-scale serving infrastructure, and the APIs developers use to bring these capabilities into production.

We’ve raised more than $125M from Lightspeed Venture Partners, Section 32, Kleiner Perkins, Microsoft’s M12 venture fund, Founders Fund, Meta, Stanford, and others. Our technology has powered experiences from companies including NVIDIA, Microsoft Xbox, Niantic, Logitech Streamlabs, Wishroll, Little Umbrella, and Bible Chat. Inworld has also been recognized by CB Insights as one of the 100 most promising AI companies globally and named one of LinkedIn’s Top 10 Startups in the USA.

About the role

Join our team as a Staff / Principal Platform Engineer and take end-to-end ownership of building, securing, and scaling our AI products. You'll be the driving force behind our cloud infrastructure, partnering with engineers across the organization to deploy and evolve services across major cloud providers using Terraform, ArgoCD, and other tooling. In this high-impact role, you'll identify what needs to be done and move it forward, directly shaping how we operate and innovate.

What you’ll do

  • Design systems for realtime, high-volume inference: plan capacity across GPU and cloud resources, build in resilience to failures, and make sure services scale smoothly as traffic grows.

  • Drive engineering velocity by identifying and building AI-powered tooling and workflows that improve how our teams develop and deploy software.

  • Facilitate a "you build it, you run it" culture by providing the necessary tools and processes for monitoring the reliability, availability, and performance of services.

  • Manage pipelines to ensure smooth and efficient code integration and deployment.

  • Conduct root cause analysis to identify critical issues and develop automated solutions to prevent recurrence.

Expected experience

  • 8-10 years of experience in software engineering.

  • 3+ years of experience with infrastructure-as-code.

  • Proficiency in managing Kubernetes clusters and applications, including creating Kustomize manifests/Helm charts for new applications.

  • Experience in creating and maintaining CI/CD pipelines for both applications and infrastructure deployments (using tools like Terraform/Terragrunt, ArgoCD, GitHub Actions, Ansible, etc.).

  • Deep knowledge of at least one major cloud provider (Google Cloud Platform, Microsoft Azure, Oracle Cloud).

  • Proficient in at least one backend programming/scripting languages such as Golang, Python, and Bash.

Candidates must be based in the SF Bay Area or willing to relocate (you will be working on-site in our South Bay office a few days a week).

The US base salary range for this full-time position is $280,000 - $350,000. In addition to base pay, total compensation includes equity and benefits. Within the range, individual pay is determined by work location, level, and additional factors, including competencies, experience, and business needs. The base pay range is subject to change and may be modified in the future.

Inworld Jobs Privacy

Skills Required

  • 8-10 years of software engineering experience
  • 3+ years of experience with infrastructure-as-code
  • Proficiency in managing Kubernetes clusters and applications
  • Experience in creating and maintaining CI/CD pipelines
  • Deep knowledge of at least one major cloud provider
  • Proficient in at least one backend programming language
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Mountain View, CA
58 Employees
Year Founded: 2021

What We Do

Inworld AI is a realtime voice AI research lab and platform. It builds the voice layer for consumer AI, the speech that powers companions, tutors, coaches, customer-service agents, and creative and content applications. Inworld's flagship product is text-to-speech, top-ranked on the independent Artificial Analysis Speech Arena, and around it the company offers a full realtime voice stack. That stack includes text-to-speech (Realtime TTS-2, with over 100 languages and natural-language voice steering), speech-to-text, and an LLM router that reaches 200+ models from every major provider with no markup. Developers can call any product on its own, for example Realtime TTS through a single API call, or combine speech-to-text, an LLM, and text-to-speech into one live conversation through the Realtime API. Dedicated GPUs are available for the largest workloads. Inworld serves developers and founders building consumer AI products. As voice quality becomes commoditized, the company's focus is realtime conversation at scale: voice agents that respond in roughly 200 milliseconds for thousands of simultaneous users, at a cost that stays efficient as usage grows. Inworld is a product-oriented research lab. Its founding team pioneered conversational AI and generative models at API.AI (acquired by Google and renamed Dialogflow), Google, and DeepMind, with expertise spanning language models, speech synthesis, multimodal interaction, and design. Inworld has raised more than $125M from investors including Lightspeed, Kleiner Perkins, Founders Fund, CRV, Intel Capital, BITKRAFT Ventures, Section 32, Meta, Microsoft's M12, and LG Technology Ventures. The company was one of six selected for the 2022 Disney Accelerator.

Similar Jobs

In-Office
La Jolla, CA, USA
121228 Employees

Pfizer Logo Pfizer

Clinical Development Medical Director, GU

Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
In-Office
7 Locations
121990 Employees
240K-400K Annually

Redfin Logo Redfin

Real Estate Agent - Tracy

Fintech • Real Estate • PropTech
In-Office
Tracy, CA, USA
5800 Employees
25K-31K Annually

Shield AI Logo Shield AI

Director, Enterprise Strategy and Operations (R5626)

Aerospace • Artificial Intelligence • Machine Learning • Robotics • Software • Defense Technology
In-Office or Remote
4 Locations
190K-290K Annually

Similar Companies Hiring

Ford Energy Thumbnail
Automotive • Software • Energy • Utilities • Manufacturing • Renewable Energy
US
55 Employees
Revel Thumbnail
Aerospace • Hardware • Robotics • Software
Marina Del Rey, California
70 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software
New York, New York
30 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account