Applied AI Engineer, Kernel Performance

Posted Yesterday
Be an Early Applicant
San Jose, CA, USA
In-Office
150K-225K Annually
Senior level
Artificial Intelligence • Hardware • Software
The Role
Build AI systems that autonomously turn new model architectures into verified, production-ready kernels optimized for Etched hardware. Develop agentic tooling to generate, compile, profile, and iterate on implementations; design evals for correctness, stability, latency, and efficiency; turn profiler traces and hardware signals into learnable datasets; ship model-generated improvements and partner with architecture teams to influence hardware-software roadmap.
Summary Generated by Built In

About Etched

Etched is building hardware for frontier intelligence. We co-design chips, racks, software, and manufacturing to deliver best-in-class throughput and latency across both prefill and decode workloads. Our first products are heavily focused on inference. Backed by hundreds of millions from top-tier investors and staffed by leading engineers, Etched is redefining the infrastructure layer for the fastest growing industry in history.

Job Summary

Every model release presents a new opportunity to push the frontier on kernel engineering. Future performance breakthroughs will come from AI systems that can understand model architectures and hardware, run thousands of experiments, learn from compiler and profiler feedback, and discover the most performant implementations faster than the best engineers.

You will build that system. Your mandate is to build AI systems that autonomously turn newly released model architectures into correct, production-ready implementations optimized for Etched hardware. These systems should explore broader design spaces, learn from every experiment, and reach peak performance faster than any traditional kernel-development workflows.

Etched offers a uniquely tight research loop: proprietary hardware, compiler, runtime, kernels, production workloads, and dedicated in-office compute under one roof. You will teach models using proprietary performance signals, iterate on their proposals, and make every experiment improve both the performance optimization system and the hardware it runs on.

Key Responsibilities

  • Own the system that turns new model architectures into verified, production-ready kernels and model mappings.

  • Build agents that understand Etched hardware, design experiments, generate implementations, compile and profile them, diagnose bottlenecks, and iterate with our teams, to the limits of model autonomy.

  • Design evals covering correctness, numerical stability, latency and efficiency.

  • Turn profiler traces, simulation, hardware counters, and expert judgment into structured signals models can learn from.

  • Curate proprietary datasets and optimization memory from complete trajectories, expert demonstrations, counterexamples, and production outcomes.

  • Build fast, reproducible experiment infrastructure and observability so experiments remain interpretable, trustworthy, and high-throughput.

  • Ship model-generated improvements to production and quantify their impact on end-to-end system performance.

  • Partner deeply with other architecture teams to shape new abstractions and Etched’s hardware-software roadmap.

  • Continuously evaluate new model releases and deploy the best for each stage of the optimization loop.

You may be a good fit if you have

  • A track record of solving hard problems across stacks and domains — you enjoy being dropped into unfamiliar territory and figuring it out

  • Comfort with both Python and low-level code: you can read it, modify it, debug it, and direct AI to write it well. We do not care whether you write code from scratch — we care whether you ship things that work.

  • Kernel experience: you've written or tuned kernels and can explain the mechanisms and performance impact of optimizations you’ve shipped

  • Fluency using AI to learn and ramp on new problems — agentic coding tools, deep research, and frontier models are how you work, not an add-on

  • Moving fluidly between research exploration, agentic experimentation, low-level debugging, and production execution.

Strong candidates may also have experience with

  • First principles thinking on accelerator performance: memory hierarchy, data movement, parallelism, synchronization, and low-precision computation.

  • Hands-on experience building and shipping LLM-based agents or AI tooling that real users depend on in production environments (beyond calling an API — context engineering, tool integration, orchestration, failure analysis)

  • An eval-driven mindset: you measure whether AI systems work before scaling them

  • Fine-tuning or post-training, RAG over proprietary data, and/or multi-agent orchestration

  • High agency and comfort with ambiguity — you find the real problem to solve

Benefits

  • Medical, dental, and vision packages with generous premium coverage

    • $500 per month credit for waiving medical benefits

  • Housing subsidy of $2k per month for those living within walking distance of the office

  • Relocation support for those moving to San Jose (Santana Row)

  • Various wellness benefits covering fitness, mental health, and more

  • Daily lunch and dinner in our office

  • Unlimited compute budget subject to ROI justification

Base Compensation Range

  • $150,000 – $225,000

How we’re different

Etched believes in the Bitter Lesson. We are the first inference-focused frontier AI system, betting early on transformer and transformer-like architectures and on increasing model sizes. Our addressable market is the entirety of inference, unlike many of our competitors.

We are a fully in-person team in San Jose (Santana Row), and greatly value engineering skills. We do not have boundaries between engineering and research, and we expect all of our technical staff to contribute to both and work across disciplines as needed.

Skills Required

  • Proficiency with Python and ability to read, modify, and debug low-level code
  • Kernel experience: written or tuned kernels and shipped performance optimizations
  • Experience using AI/LLMs and agentic tools to learn and ramp on new problems
  • Ability to build systems that generate implementations, compile, profile, diagnose bottlenecks, and iterate
  • Experience designing evaluations for correctness, numerical stability, latency, and efficiency
  • Experience turning profiler traces, simulation, and hardware counters into structured signals for models
  • Comfort operating across research exploration, low-level debugging, and production execution
  • First-principles understanding of accelerator performance (memory hierarchy, data movement, parallelism)
  • Hands-on experience building and shipping LLM-based agents or AI tooling in production
  • Eval-driven mindset and experience with fine-tuning, RAG, or multi-agent orchestration
  • High agency and comfort with ambiguity; ability to find the real problem to solve

Etched Compensation & Benefits Highlights

The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about Etched and has not been reviewed or approved by Etched.

  • Equity Value & Accessibility Equity growth is described as strong and significant equity is part of the package. High total compensation for technical roles reinforces the equity-led upside.
  • Healthcare Strength Medical, dental, and vision coverage include generous premium support, indicating robust core healthcare. This reduces employee cost exposure for essential coverage.
  • Wellbeing & Lifestyle Benefits Daily lunch and dinner, a housing subsidy for those living near the office, relocation support, and wellness perks are highlighted. These offerings lower day-to-day living costs and support practical wellbeing.

Etched Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Pakenham
53 Employees
Year Founded: 2022

What We Do

By burning the transformer architecture into our chips, we’re creating the world’s most powerful servers for transformer inference.

Similar Jobs

Notion Logo Notion

Strategic Finance, Business Partnership & Workflow Automation

Artificial Intelligence • Productivity • Software
Hybrid
San Francisco, CA, USA
1000 Employees
150K-165K Annually

3Play Media Logo 3Play Media

Live Voice Writer (English) (Contractor)

Artificial Intelligence • Information Technology • Professional Services • Social Impact • Software
Easy Apply
Remote or Hybrid
U.S.
211 Employees
1-1 Hourly

Golden Hippo Logo Golden Hippo

Digital Acquisition Marketing Intern

Digital Media • eCommerce • Information Technology • Marketing Tech • Retail • Social Media • Analytics
Remote or Hybrid
United States
500 Employees
17-20 Hourly

Golden Pet Brands Logo Golden Pet Brands

Production Assistant Intern

Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
Easy Apply
In-Office
2 Locations
178 Employees
17-20 Hourly

Similar Companies Hiring

Hanover Park Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
42 Employees
Kepler  Thumbnail
Fintech • Software
New York, New York
6 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account