Fieldguide is establishing a new state of trust for global commerce and capital markets by automating and streamlining the work of assurance and audit practitioners—specifically in cybersecurity, privacy, and financial audits. We build software for the people who enable trust between businesses.
We're based in San Francisco, CA, and backed by Goldman Sachs Alternatives, Bessemer Venture Partners, 8VC, Floodgate, Y Combinator, and more. Over 50 of the top 100 accounting and consulting firms trust Fieldguide to power mission-critical work.
The Foundation Agents team stewards the long-horizon agents powering the Fieldguide AI platform. We work at the frontier of AI product development: agent knowledge, evaluations, and improving quality and reliability at scale. As a Senior Software Engineer, Agents, you'll take ownership of how the team measures and improves agent quality, and help drive the platform forward.
Evals strategy and error-analysis practice, shaping how the team measures and improves agent quality
Design and build agent knowledge and evaluation infrastructure for Fieldguide's long-horizon agents
Lead error analysis on agent behavior, turning findings into concrete platform-level reliability improvements
Build and harden backend systems that support agent execution, evaluation, and monitoring at scale
Drive the AI platform's reliability and quality roadmap forward, working closely with the broader AI team
Mentor engineers on the team, raising the bar on eval rigor and error-analysis practice
You've built AI products end-to-end, with real ownership over agent quality outcomes
You think in evals and error analysis as a discipline; you dig into why an agent failed and fix the systemic cause
You're strong in the backend and comfortable owning platform-level systems
You're motivated by long-horizon agents that do real work in production
You multiply the people around you, not just your own output
Must-have:
1+ years working specifically on agents
Strong experience working on an AI platform
Demonstrated experience building evals and performing error analysis
Backend engineering experience
Nice-to-have:
Strong platform engineering skills
Frontend experience
Distributed systems experience
Long-horizon agents: Working on agents that do real, sustained work
Evaluation as a craft: Evals and error analysis are core to how this team improves quality, not an afterthought
Platform-level impact: Your work shapes the reliability and quality of every agent built on top of it
Frontier problems: You're working on open problems in agent reliability that don't have established playbooks yet
Competitive compensation with equity
Comprehensive health and wellness benefits
Flexible time off and work schedules
Technology reimbursements
401(k) plan
Twice-yearly in-person offsites across the U.S.
Wellness benefits starting on your first day
Fearless — Inspire and break down seemingly impossible walls
Fast — Launch fast with excellence; iterate to perfection
Lovable — Deliver happiness and 11-star experiences
Owners — Execute and run the business with ownership
Win-win — Create mutual value and earn trust for life
Inclusive — Scale the best ideas with inclusive teams
Skills Required
- At least 1 year of experience working specifically on AI agents
- Strong experience working on an AI platform
- Demonstrated experience building evaluations and performing error analysis
- Backend engineering experience
- Strong platform engineering skills
- Frontend experience
- Distributed systems experience
Fieldguide Compensation & Benefits Highlights
-
Healthcare Strength — Health coverage is described as comprehensive across medical, dental, vision, and mental health, with wellness benefits that start day one and a bundle of free therapy sessions on some roles. HSA/FSA options and life/disability insurance are also included in the package.
-
Leave & Time Off Breadth — Flexible PTO and flexible work schedules are emphasized, with paid holidays and sick leave noted alongside remote or hybrid options depending on the role. Twice‑yearly in‑person offsites complement the remote setup.
-
Equity Value & Accessibility — Meaningful equity/stock options are positioned as a core component of total rewards and appear broadly included across postings. Some roles also indicate performance bonuses in addition to competitive cash compensation.
Fieldguide Insights
What We Do
Build the expertise the rest of the industry will follow. Fieldguide is establishing a new state of trust for global commerce and capital markets by automating and streamlining the work of assurance and audit practitioners. We build agentic software for the people who enable trust between businesses. This is where deep professional expertise meets cutting-edge AI. Our team is reimagining how the world's trust is earned, and we're building the people who define how AI works in a domain with real stakes and high data complexity. If you want a career the rest of your industry hasn't caught up to yet, you'll find it here. Backed by top investors including Growth Equity at Goldman Sachs Alternatives, Bessemer Venture Partners, 8VC, Floodgate, and Y Combinator, we are on a growth journey few companies realize.
Why Work With Us
You'll matter here. We're an early-stage company where your work has outsized impact. We've already launched Field Agents, agentic AI that transforms audit and advisory workflows, and we're just getting started. What you build here will be followed by the rest of the industry.
Fieldguide Offices
Hybrid Workspace
Employees engage in a combination of remote and on-site work.