Staff / Principal Platform Engineer - USA

Reposted 27 Days Ago
Be an Early Applicant
Mountain View, CA, USA
Hybrid
280K-350K Annually
Senior level
Software
The Role
As a Staff/Principal Platform Engineer, you will build, secure, and scale AI products, manage cloud infrastructure, integrate CI/CD pipelines, and enhance developer workflows with automated tooling.
Summary Generated by Built In

About Inworld

Inworld is a research lab and inference provider focused on realtime AI for consumer-facing applications. We build first-party speech models, serve LLMs, and run the inference behind modular APIs designed for high-volume, realtime workloads.

Hundreds of millions of users interact with Inworld powered apps every day and we serve over 10 trillion LLM tokens per month. Our models and infrastructure support consumer applications across companions, healthcare, fitness, education, media, and more. Our work spans model research, realtime inference, large-scale serving infrastructure, and the APIs developers use to bring these capabilities into production.

We’ve raised more than $125M from Lightspeed Venture Partners, Section 32, Kleiner Perkins, Microsoft’s M12 venture fund, Founders Fund, Meta, Stanford, and others. Our technology has powered experiences from companies including NVIDIA, Microsoft Xbox, Niantic, Logitech Streamlabs, Wishroll, Little Umbrella, and Bible Chat. Inworld has also been recognized by CB Insights as one of the 100 most promising AI companies globally and named one of LinkedIn’s Top 10 Startups in the USA.

About the role

Join our team as a Staff / Principal Platform Engineer and take end-to-end ownership of building, securing, and scaling our AI products. You'll be the driving force behind our cloud infrastructure, partnering with engineers across the organization to deploy and evolve services across major cloud providers using Terraform, ArgoCD, and other tooling. In this high-impact role, you'll identify what needs to be done and move it forward, directly shaping how we operate and innovate.

What you’ll do

  • Work closely with engineers to design, deploy, and maintain reliable, high-performance, and secure cloud infrastructure for our TTS and LLM Router.

  • Drive engineering velocity by identifying and building AI-powered tooling and workflows that improve how our teams develop and deploy software.

  • Facilitate a "you build it, you run it" culture by providing the necessary tools and processes for monitoring the reliability, availability, and performance of services.

  • Manage pipelines to ensure smooth and efficient code integration and deployment.

  • Conduct root cause analysis to identify critical issues and develop automated solutions to prevent recurrence.

Expected experience

  • 8-10 years of experience in software engineering.

  • 3+ years of experience with infrastructure-as-code.

  • Proficiency in managing Kubernetes clusters and applications, including creating Kustomize manifests/Helm charts for new applications.

  • Experience in creating and maintaining CI/CD pipelines for both applications and infrastructure deployments (using tools like Terraform/Terragrunt, ArgoCD, GitHub Actions, Ansible, etc.).

  • Deep knowledge of at least one major cloud provider (Google Cloud Platform, Microsoft Azure, Oracle Cloud).

  • Proficient in at least one backend programming/scripting languages such as Golang, Python, and Bash.

Candidates must be based in the SF Bay Area or willing to relocate (you will be working on-site in our South Bay office a few days a week).

The US base salary range for this full-time position is $280,000 - $350,000. In addition to base pay, total compensation includes equity and benefits. Within the range, individual pay is determined by work location, level, and additional factors, including competencies, experience, and business needs. The base pay range is subject to change and may be modified in the future.

Inworld Jobs Privacy

Skills Required

  • 8-10 years of software engineering experience
  • 3+ years of experience with infrastructure-as-code
  • Proficiency in managing Kubernetes clusters and applications
  • Experience in creating and maintaining CI/CD pipelines
  • Deep knowledge of at least one major cloud provider
  • Proficient in at least one backend programming language
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Mountain View, CA
58 Employees
Year Founded: 2021

What We Do

Inworld AI is a realtime voice AI research lab and platform. It builds the voice layer for consumer AI, the speech that powers companions, tutors, coaches, customer-service agents, and creative and content applications. Inworld's flagship product is text-to-speech, top-ranked on the independent Artificial Analysis Speech Arena, and around it the company offers a full realtime voice stack. That stack includes text-to-speech (Realtime TTS-2, with over 100 languages and natural-language voice steering), speech-to-text, and an LLM router that reaches 200+ models from every major provider with no markup. Developers can call any product on its own, for example Realtime TTS through a single API call, or combine speech-to-text, an LLM, and text-to-speech into one live conversation through the Realtime API. Dedicated GPUs are available for the largest workloads. Inworld serves developers and founders building consumer AI products. As voice quality becomes commoditized, the company's focus is realtime conversation at scale: voice agents that respond in roughly 200 milliseconds for thousands of simultaneous users, at a cost that stays efficient as usage grows. Inworld is a product-oriented research lab. Its founding team pioneered conversational AI and generative models at API.AI (acquired by Google and renamed Dialogflow), Google, and DeepMind, with expertise spanning language models, speech synthesis, multimodal interaction, and design. Inworld has raised more than $125M from investors including Lightspeed, Kleiner Perkins, Founders Fund, CRV, Intel Capital, BITKRAFT Ventures, Section 32, Meta, Microsoft's M12, and LG Technology Ventures. The company was one of six selected for the 2022 Disney Accelerator.

Similar Jobs

Comcast Logo Comcast

Senior Solutions Architect

Digital Media • Information Technology • News + Entertainment
Hybrid
Irvine, CA, USA
115000 Employees
174K-272K Annually

PwC Logo PwC

Pricing and Revenue Consulting Manager - Consumer Markets Sector

Artificial Intelligence • Professional Services • Business Intelligence • Consulting • Cybersecurity • Generative AI
Hybrid
58 Locations
370000 Employees
99K-232K Annually

PwC Logo PwC

Sanctions - Crypto & Digital Assets - Senior Manager

Artificial Intelligence • Professional Services • Business Intelligence • Consulting • Cybersecurity • Generative AI
Hybrid
8 Locations
370000 Employees
124K-280K Annually

PwC Logo PwC

Assurance Internal Communications Director

Artificial Intelligence • Professional Services • Business Intelligence • Consulting • Cybersecurity • Generative AI
Hybrid
69 Locations
370000 Employees
123K-123K Annually

Similar Companies Hiring

Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees
Revel.io Thumbnail
Aerospace • Hardware • Robotics • Software
US
50 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software
New York, New York
30 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account