Senior Machine Learning Engineer, ML Infrastructure- Online

Posted Yesterday
Hiring Remotely in Seattle, WA, USA
Remote or Hybrid
187K-243K Annually
Senior level
AdTech • Artificial Intelligence • Gaming • Machine Learning • Software • Virtual Reality • Metaverse
Unity is the leading platform to create and grow games and interactive experiences.
The Role
Design, build, and operate large-scale online ML inference infrastructure to serve production models with low latency and high reliability. Optimize inference performance, enable distributed training workflows, integrate ML pipelines with orchestration systems, improve observability, and support safe model rollout, canary testing, and automated rollback. Lead architectural improvements and collaborate with ML engineers and product teams to scale, monitor, and streamline model deployment and iteration.
Summary Generated by Built In

The Role

We are seeking a Senior ML engineer to design and evolve Unity Vector’s online model inference platform. This role focuses on building reliable infrastructure for serving machine learning models in production, optimizing inference performance, and enabling safe, efficient experimentation across high-traffic online systems.

You will work closely with ML engineers, platform teams, and product stakeholders to ensure models can be deployed, scaled, monitored, and iterated on efficiently. You will play a key role in shaping how models are packaged, served, validated, monitored, and optimized in production environments.

This role requires strong systems thinking, deep experience with production ML infrastructure, and the ability to drive architectural improvements across teams.

What you'll be doing

  •  Design and operate large-scale online inference infrastructure that serves production ML models with low latency and high reliability, such as PyTorch, Triton Inference Server, Kubernetes, GKE, Ray, or similar distributed serving frameworks.

  • Develop infrastructure that supports distributed training workflows using technologies such as Pytorch, Ray Data, and Ray Train, etc.

  • Integrate ML pipelines with workflow orchestration systems (e.g., Flyte, Airflow, or similar) to enable reliable multi-stage training workflows

  • Optimize model performance through model compilation, GPU/CPU utilization improvements, request scheduling, kernel fusion, and runtime-level tuning.

  • Improve observability of ML systems through latency, throughput, error-rate, cost, saturation, and model-health monitoring.

  • Partner closely with ML engineers to support faster model iteration while maintaining production safety, scalability, and cost efficiency.

  • Improve the reliability and reproducibility of model serving workflows, including model packaging, artifact validation, compatibility testing, and deployment automation.

  • Lead architectural improvements that make the online ML platform more robust, user-friendly, scalable, and cost-efficient. 

What we're looking for

  •  Experience building and operating production-grade online ML inference systems, such as NVIDIA Triton Inference Server, TorchServe, Ray Serve, TensorFlow Serving, or similar systems.

  • Experience with model serving frameworks such as NVIDIA Triton Inference Server, TorchServe, Ray Serve, TensorFlow Serving, or similar systems.

  •  Experience optimizing inference workloads using techniques such as dynamic batching, model compilation, quantization, GPU acceleration, GPU kernel optimization, caching, or runtime tuning. 

  •  Strong experience with distributed systems, Kubernetes, autoscaling, service reliability, and production observability.

  • Strong programming skills in Python, with practical experience working on production ML systems and high-scale services. 

  • Experience with PyTorch and modern model deployment workflows, including model packaging, validation, and serving lifecycle management.

  • Experience designing infrastructure for safe model rollout, canary testing, A/B experimentation, and automated rollback.

  • Strong systems thinking, with the ability to reason about latency, throughput, reliability, scalability, and cost tradeoffs in online systems.

  • Proven ability to lead technical direction and influence architectural decisions across teams without formal authority. 

Additional information

  • Relocation support is not available for this position

  • Work visa/immigration sponsorship is not available for this position

  • $187,200$-$243,300


This range reflects the anticipated base salary for this position. Beyond base salary, this role may be eligible for equity awards and participation in our company incentive plans (such as annual discretionary bonuses or sales commissions). The final offer amount will depend on several factors, including geographic location and the candidate’s relevant experience, professional background, and skill set.


Benefits


At Unity, we want our team members to thrive. We offer a wide range of benefits designed to support well-being and work-life balance.


Please note: Benefits eligibility, specific offerings, and coverage vary based on the country and employment status.


While specific benefits vary, here are some of the ways we strive to take care of our eligible team members globally: Comprehensive health, life, and disability insurance | Commute subsidy | Employee stock ownership | Competitive retirement/pension plans | Generous vacation and personal days | Support for new parents through leave and family-care programs | Office food snacks | Mental Health and Wellbeing programs and support | Employee Resource Groups | Global Employee Assistance Program | Training and development programs | Volunteering and donation matching program


Life at Unity


Unity [NYSE: U] is the world’s leading game engine, powering play for more than 3 billion consumers each month. The top mobile games in the world, the most played PC indie titles, the most innovative console games, and virtually all of the top XR and Web Games are developed, deployed, and grown in Unity. Unity also enables teams across industries like automotive, manufacturing, and healthcare to design, simulate, and collaborate in 3D — closing the gap between ideas and reality. For more information, please visit www.unity.com.


Unity is a proud equal opportunity employer. We are committed to fostering an inclusive, innovative environment and celebrate our employees across age, race, color, ancestry, national origin, religion, disability, sex, gender identity or expression, sexual orientation, or any other protected status in accordance with applicable law. Our differences are strengths that enable us to support the growing and evolving needs of our customers, partners, and collaborators. If you have a disability that means there are preparations or accommodations we can make to help ensure you have a comfortable and positive interview experience, please fill out this form to let us know.


This position requires the incumbent to have a sufficient knowledge of English to have professional verbal and written exchanges in this language since the performance of the duties related to this position requires frequent and regular communication with colleagues and partners located worldwide and whose common language is English.
This posting is intended to fill an existing vacancy, and we are committed to providing applicants with updates throughout the hiring process in accordance with applicable law.


Headhunters and recruitment agencies may not submit resumes/CVs through this website or directly to managers. Unity does not accept unsolicited headhunter and agency resumes. Unity will not pay fees to any third-party agency or company that does not have a signed agreement with Unity.
Your privacy is important to us. Please take a moment to review our
Prospect and Applicant Privacy Policies. Should you have any concerns about your privacy, please contact us at [email protected].

Skills Required

  • Experience building and operating production-grade online ML inference systems (e.g., Triton, TorchServe, Ray Serve, TensorFlow Serving).
  • Strong programming skills in Python and production ML system development.
  • Experience with PyTorch and modern model deployment workflows, including model packaging, validation, and serving lifecycle management.
  • Experience optimizing inference workloads using dynamic batching, model compilation, quantization, GPU acceleration, kernel optimization, caching, or runtime tuning.
  • Strong experience with distributed systems, Kubernetes (GKE), autoscaling, and service reliability.
  • Experience with Ray, Ray Data, Ray Train for distributed training and Ray Serve for serving.
  • Experience integrating ML pipelines with workflow orchestration systems (e.g., Flyte, Airflow) for reliable multi-stage training workflows.
  • Experience designing infrastructure for safe model rollout, canary testing, A/B experimentation, and automated rollback.
  • Improve observability of ML systems (latency, throughput, error-rate, cost, saturation, model-health monitoring).
  • Proven ability to lead technical direction and influence architectural decisions across teams.
  • Strong systems thinking with ability to reason about latency, throughput, reliability, scalability, and cost tradeoffs.

Unity Compensation & Benefits Highlights

The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about Unity and has not been reviewed or approved by Unity.

  • Healthcare Strength Core medical, dental, vision, life/disability, and mental‑health/EAP offerings are positioned as comprehensive across eligible locations. This breadth aligns with large‑tech standards and is highlighted in official materials.
  • Retirement Support A 401(k) plan with employer matching is part of the U.S. package. Retirement benefits are characterized as competitive and a stable element of total rewards.
  • Parental & Family Support Paid parental leave and family‑care support are emphasized, with indications of generous time off for new parents. These programs are presented as global in scope, with specifics verified by location.

Unity Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: San Francisco, CA
4,500 Employees
Year Founded: 2004

What We Do

Unity [NYSE: U] is the world’s leading game engine, powering play for more than 3 billion consumers each month. The top mobile games in the world, the most played PC indie titles, the most innovative console games, and virtually all of the top XR and Web Games are developed, deployed, and grown in Unity. Unity also enables teams across industries like automotive, manufacturing, and healthcare to design, simulate, and collaborate in 3D — closing the gap between ideas and reality.

Why Work With Us

We believe the world is a better place with more creators in it. This is at the core of our business because we believe our technology can change the world. Our products give content creators the tools to not just entertain but to create innovative RT3D experiences and deliver better processes for almost every industry.

Gallery

Gallery

Similar Jobs

Remote
United States
101 Employees
120K-195K Annually

BJAK Logo BJAK

Senior Machine Learning Engineer

Artificial Intelligence • Fintech • Software • Financial Services
Remote or Hybrid
United States
253 Employees

Affirm Logo Affirm

Security Engineer

Big Data • Fintech • Mobile • Payments • Financial Services
Easy Apply
Remote
United States
2200 Employees
204K-290K Annually

Shield AI Logo Shield AI

Analytics Engineer

Aerospace • Artificial Intelligence • Machine Learning • Robotics • Software
Remote
USA
120K-180K Annually

Similar Companies Hiring

Hanover Park Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
42 Employees
Kepler  Thumbnail
Fintech • Software
New York, New York
6 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account