Applied AI Engineer

Reposted One Month Ago
San Francisco, CA, USA
In-Office
180K-240K Annually
Mid level
Artificial Intelligence • Enterprise Web • Productivity • Software
The Role
Build, refine, and scale production agent systems: multi-step tool-using agents, memory systems, proactive agents, and evaluation frameworks. Ensure fast, predictable, traceable behavior and uptime while working across infra, frontend, and product.
Summary Generated by Built In
⚠️ Please read first
  • This is a full-time, in-person role based in San Francisco (Presidio) - we work from the office 5 days a week.

  • You must be based in the Bay Area or willing to relocate before starting.

  • We require US work authorisation, but are open to O-1 or J-1 visa sponsorship for exceptional candidates.

About Sauna by Wordware

Every tool you use was built as an island, so you became the glue between them: a name from Gmail into the CRM, a time from the CRM into Calendly, a note from Calendly into a doc. AI showed up and mostly became an eleventh window to paste into.

Sauna is a hybrid of AI and software standing on one shared brain. It holds your context and hands you real software built on top of it, so an ordinary Calendly becomes one that writes the invite itself and knows which of your busy hours bend, and for whom. Software for the path you repeat, AI for the edge that never repeats.

Wordware is the team behind it, 15 people backed by Spark Capital, Felicis, and Y Combinator with a $30M seed, the largest in YC's history. We build from an office in the Presidio, fifty metres from Crissy Field beach, and we work absurdly hard because we can see and feel the outcome every day.

About the Role:

As an Applied AI Engineer, you’ll be responsible for building, refining, and scaling the agent systems inside Sauna, from architecture to evals to deployment.

We care about what works in production: fast response times, predictable behavior, traceability, and uptime.

You’ll work across infra, frontend, and product to make sure the agents people build inside Wordware actually work.

A few examples of what you might work on:
  • Implement multi-step, tool-using agents that hit real APIs and handle retries, auth, timeouts, and edge cases.

  • Design agent memory systems that persist relevant state across runs, e.g. memory migrations, context organization, and orchestration state.

  • Create agents that proactively do work and send you reminders.

  • Own and evolve our eval framework: both automated checks and human-in-the-loop scoring.

  • Plus whatever else you see fit.

Who You AreMinimum
  • 3+ years of engineering experience, including time shipping production software.

  • You've built and deployed agent-like systems: multi-step LLM pipelines, tool-using bots, scripted assistants, or similar.

  • Hands-on experience with:

    • Agent orchestration and memory management (e.g. memory migrations, state organization).

    • Tool use and orchestration (e.g. calling real APIs, using plugins, auth flows)

    • Evaluation: success metrics, regression testing, and improving agent behavior over time

  • You write production-grade code and can work across systems without needing a spec.

  • You'd rather ship than polish forever.

Bonus (not required)
  • Shipped agents that live in the wild, used by customers, not just internal demos.

  • Familiarity with LLM ops, tracing, observability, and failure handling.

  • You've been a founder or early engineer, and it shows in the bar you hold your own work to.

Compensation & Benefits:

Base salary: $180K–$240K + meaningful early-stage equity + health, dental, 401(k), considerable PTO, gym budget, lunch.

The Process

We keep our process simple. Exceptional candidates go from first touch to offer within 2 weeks.

  1. Application: Submit your resume and answer a few quick questions.

  2. 15-min intro call: Quick check to align on location, motivation, and logistics. If it’s a go, we move fast from here.

  3. System design interview (1 hour): We dig into how you think about agent design: architecture, tradeoffs, and your experience building AI harnesses and agent systems.

  4. Technical interview (optional follow-up): A coding round testing hands-on engineering fluency and speed, only if we need a closer look.

  5. Final conversation: Answer any questions and scope out the work trial.

  6. Work trial: Paid, in-person. Typically 5 days, though we're flexible on timing depending on the role. You’ll work on something meaningful with us.

Skills Required

  • 3+ years engineering experience including shipping production software
  • Built and deployed agent-like systems (multi-step LLM pipelines, tool-using bots, scripted assistants)
  • Hands-on experience with agent orchestration and memory management (state organization, memory migrations)
  • Hands-on experience with tool use and orchestration (calling real APIs, plugins, auth flows, retries, timeouts)
  • Experience designing evaluation frameworks: success metrics, regression testing, human-in-the-loop scoring
  • Ability to write production-grade code and work across systems without needing a spec
  • Full-time, in-person role based in San Francisco (Presidio), 5 days/week; must be Bay Area–based or relocate before starting
  • US work authorization (company open to O-1 or J-1 sponsorship for exceptional candidates)
  • Shipped customer-facing agents (not just internal demos)
  • Familiarity with LLM ops, tracing, observability, and failure handling
  • Experience as a founder or early engineer
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
18 Employees

What We Do

Sauna is an AI agent designed to learn how individuals and teams work, remembering critical information and acting on their behalf. By connecting to tools like email, calendar, and messaging, it helps manage tasks and workflows automatically, allowing work to move forward even when the user is offline. It effectively serves as an AI-powered workspace and coworkers' assistant.

Similar Jobs

AKASA Logo AKASA

Software Engineer

Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Software • Generative AI
Hybrid
3 Locations
110 Employees
180K-250K Annually

Deffai Logo Deffai

Full-stack Engineer

Artificial Intelligence • Software • Biotech • Pharmaceutical
In-Office or Remote
San Francisco, CA, United States
8 Employees
180K-260K Annually

Goaly Logo Goaly

Artificial Intelligence Engineer

Artificial Intelligence • Machine Learning • Software • Generative AI
In-Office
Palo Alto, CA, USA
80 Employees
In-Office
26 Locations
456553 Employees
68K-219K Annually

Similar Companies Hiring

Revel Thumbnail
Aerospace • Hardware • Robotics • Software
Marina Del Rey, California
60 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software
New York, New York
30 Employees
Vega Thumbnail
Artificial Intelligence • Automotive • Insurance • Transportation
US
43 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account