AI Software Engineer lll

Sorry, this job was removed at 10:47 p.m. (UTC) on Thursday, Sep 03, 2026
Kirkland, WA, USA
Hybrid
152K-205K Annually
Mid level
Computer Vision • eCommerce • Healthtech • Internet of Things • Security
Making great technology accessible!
The Role
Build and operate infrastructure for production AI systems across model training, experimentation, deployment, inference, observability, and developer tooling. Design scalable distributed systems, cloud services, containerized workloads, and AI platform capabilities. Improve reliability, performance, cost efficiency, and operational workflows while supporting multimodal workloads. Diagnose issues across application and infrastructure layers, evaluate emerging AI technologies, and own systems from architecture through continuous production improvement.
Summary Generated by Built In
At Wyze, we make smart home technology accessible to everyone. We're known for disrupting markets with high-quality, affordable products - from cameras to lighting to sensors and more. We believe technology should simplify life, not complicate it. We’re a fast-moving, customer-obsessed team driven by curiosity and powered by data. 
We are looking for a Software Engineer to help build and scale the infrastructure behind our production AI systems. You will work on the platform that supports the full lifecycle of our AI models, from training and experimentation to deployment and production inference. This includes building scalable model training infrastructure, production AI services, cloud and compute infrastructure, deployment systems, observability, and developer tooling.

This is primarily an infrastructure and systems engineering role rather than a model research role. You do not need to be an expert in LLM inference or GPU optimization when you join. We are looking for a strong software engineer who can build reliable distributed systems, learn quickly, and solve infrastructure problems across different layers of the stack. You will work closely with AI scientiests and other software engineers to make it easier and faster to train, deploy, operate, and iterate on AI models in production. The AI infrastructure landscape is evolving extremely quickly. We value engineers who can evaluate new technologies pragmatically, move fast, adapt to changing requirements, and continuously improve how we build AI systems.

What You’ll Work On
  • Design, build, and operate infrastructure and backend services that power production AI features.
  • Build and improve model training infrastructure, including systems that support training jobs, experimentation, compute management, data workflows, and model artifacts.
  • Build infrastructure that supports the full model lifecycle, from training and experimentation through deployment and production serving.
  • Improve the scalability, performance, reliability, and cost efficiency of our AI platform.
  • Build and maintain cloud-based services and containerized workloads for AI and ML applications.
  • Develop systems that make it easier for ML engineers to train, evaluate, deploy, and iterate on models.
  • Build reusable platform capabilities that allow engineering teams to launch new AI-powered features quickly and safely.
  • Improve deployment, rollout, monitoring, and operational workflows for AI models and services.
  • Diagnose and resolve reliability and performance issues across application, infrastructure, compute, and ML system layers.
  • Support large-scale AI workloads across text, image, video, and multimodal applications.
  • Evaluate and adopt emerging AI infrastructure technologies when they provide meaningful improvements in productivity, performance, reliability, or cost.
  • Improve software development and operational workflows through effective use of modern AI coding agents and agentic engineering tools.
  • Work closely with AI scientiests, and product teams to translate rapidly changing requirements into practical technical solutions.
  • Own systems end-to-end, from architecture and implementation through deployment, monitoring, operation, and continuous improvement.

Required Qualifications
  • 3+ years of professional software engineering experience building production backend systems, infrastructure, or distributed systems.
  • Strong programming skills in Python, Java, Go, or a comparable backend or systems language.
  • Strong understanding of distributed systems fundamentals, including concurrency, fault tolerance, messaging, backpressure, load balancing, and horizontal scaling.
  • Production experience with Kubernetes or Docker-based containerized workloads.
  • Experience operating production services on a major cloud platform such as AWS, GCP, or Azure.
  • Experience designing and operating high-throughput or latency-sensitive production systems.
  • Familiarity with NoSQL databases such as DynamoDB, Cassandra, or comparable distributed data stores, including common data modeling and scalability considerations.
  • Strong debugging and problem-solving skills across application, infrastructure, and networking layers.
  • Experience with production observability, including metrics, logging, tracing, dashboards, and alerting.
  • Deep understanding of state-of-the-art AI coding agents, such as Claude Code, OpenAI Codex, or comparable agentic coding systems, with demonstrated ability to use them effectively in real software engineering workflows.
  • Ability to independently own complex systems from design through production operations.
  • Ability to move quickly, iterate with incomplete information, and adapt effectively as product requirements and technical priorities evolve.
  • Comfort operating in a fast-paced, high-pressure environment while maintaining sound engineering judgment and execution quality.

Preferred Qualifications
  • Understanding of LLM inference architecture and production model-serving systems.
  • Experience with inference frameworks such as SGLang, vLLM, TensorRT-LLM, TGI, or similar technologies.
  • Understanding of inference concepts such as prefill vs. decode, continuous batching, KV cache, prefix caching, and request scheduling.
  • Experience with NVIDIA GPUs such as A100 or H100.
  • Experience profiling or optimizing GPU workloads for throughput, latency, memory utilization, or cost.
  • Experience with Kubernetes GPU scheduling, GPU node pools, KEDA, or workload-aware autoscaling.
  • Experience with high-throughput image, video, or multimodal processing systems.
Compensation
The base pay range for this role is $152,000 – $205,000 per year.

Skills Required

  • 3+ years of professional software engineering experience building production backend systems, infrastructure, or distributed systems
  • Strong programming skills in Python, Java, Go, or a comparable backend or systems language
  • Strong understanding of distributed systems fundamentals, including concurrency, fault tolerance, messaging, backpressure, load balancing, and horizontal scaling
  • Production experience with Kubernetes or Docker-based containerized workloads
  • Experience operating production services on AWS, GCP, Azure, or another major cloud platform
  • Experience designing and operating high-throughput or latency-sensitive production systems
  • Familiarity with NoSQL databases such as DynamoDB, Cassandra, or comparable distributed data stores
  • Strong debugging and problem-solving skills across application, infrastructure, and networking layers
  • Experience with production observability, including metrics, logging, tracing, dashboards, and alerting
  • Deep understanding of AI coding agents such as Claude Code, OpenAI Codex, or comparable agentic coding systems, with demonstrated effective use in software engineering workflows
  • Ability to independently own complex systems from design through production operations
  • Ability to move quickly, iterate with incomplete information, and adapt to changing product requirements and technical priorities
  • Comfort operating in a fast-paced, high-pressure environment while maintaining sound engineering judgment and execution quality
  • Understanding of LLM inference architecture and production model-serving systems
  • Experience with inference frameworks such as SGLang, vLLM, TensorRT-LLM, TGI, or similar technologies
  • Understanding of prefill versus decode, continuous batching, KV cache, prefix caching, and request scheduling
  • Experience with NVIDIA GPUs such as A100 or H100
  • Experience profiling or optimizing GPU workloads for throughput, latency, memory utilization, or cost
  • Experience with Kubernetes GPU scheduling, GPU node pools, KEDA, or workload-aware autoscaling
  • Experience with high-throughput image, video, or multimodal processing systems

Similar Jobs

Square Logo Square

Account Executive

eCommerce • Fintech • Hardware • Payments • Software • Financial Services
Remote or Hybrid
Everett, WA, USA
12000 Employees
129K-233K Annually

Expedia Group Logo Expedia Group

Senior Product Manager

AdTech • eCommerce • Information Technology • Software • Travel • Generative AI
Hybrid
Seattle, WA, USA
16000 Employees
173K-277K Annually

Expedia Group Logo Expedia Group

Application Security Engineer

AdTech • eCommerce • Information Technology • Software • Travel • Generative AI
Hybrid
Seattle, WA, USA
16000 Employees
146K-234K Annually

Expedia Group Logo Expedia Group

Development Engineer

AdTech • eCommerce • Information Technology • Software • Travel • Generative AI
Hybrid
Seattle, WA, USA
16000 Employees
146K-234K Annually
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Kirkland, WA
350 Employees
Year Founded: 2017

What We Do

It’s our goal to become the most user-centric smart home technology company. We’re passionate about providing users access to high-quality products at great prices, we relentlessly keep costs low by partnering with the world’s most efficient manufacturers, we cut out “channel fat” by selling directly from our own website, and, unlike our competitors, we don’t seek a high-profit margin over our cost base, passing on all of these savings to our users. As we grow, we will continue to launch high-quality, affordable smart home products that enrich people’s lives and make great technology accessible to everyone!

Why Work With Us

We’re passionate about providing customers access to high-quality products at great prices. We relentlessly keep costs low by partnering with the world’s most efficient manufacturers. We cut out “channel fat” by selling directly from our own website.

Gallery

Gallery

Similar Companies Hiring

Milestone Systems Thumbnail
Artificial Intelligence • Security • Software • Analytics • Big Data Analytics
Lake Oswego, OR
1500 Employees
OneImaging Thumbnail
Healthtech
Miami, FL
62 Employees
Golden Pet Brands Thumbnail
Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
El Segundo, California
178 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account