Technical Program Manager - Embedded Evaluations and Investigations

Posted 5 Days Ago
Be an Early Applicant
San Francisco, CA, USA
In-Office
250K-350K Annually
Entry level
Artificial Intelligence • Software
The Role
Own end-to-end technical engagements involving AI evaluations, embedded assessments, and rapid-response incident analysis. Scope projects, coordinate technical teams, manage dependencies, guide methodological decisions, maintain partner relationships, interpret evaluation results, and present findings to technical, governmental, media, and public audiences. The role also develops repeatable workflows and playbooks while collaborating with research and governance teams.
Summary Generated by Built In
Salary range: $250,000 - $350,000/year + benefits + bonus

Description: Transluce is a fast-moving nonprofit research lab building the public tech stack for AI evaluation and oversight. We partner with frontier AI developers, governments, and civil society organizations to evaluate how AI systems behave, from misaligned agent swarms to child safety and mental health.

About the role: As a Technical Program Manager, you will scope, organize, and direct the work of our technical staff across our external engagements. These include embedded evaluations on site at frontier labs, rapid-response incident analysis, and evaluation services for partners such as the EU AI Office and Common Sense Media. You will own each engagement end to end, from the first scoping conversation to the final dissemination.
As an early member of a highly collaborative team, you will learn and grow quickly, and work closely with our technical and governance teams to protect researcher time and turn one-off engagements into repeatable playbooks.

Core responsibilities:
  • Define the scope of each engagement, set key milestones, and negotiate terms, with support from our operations team.
  • Organize the technical team's work, manage dependencies, and keep everyone aligned through frequent touchpoints.
  • Act as a sounding board on key strategic and methodological questions, reflecting the needs of the engagement and other considerations, such as relevant regulatory guidance.
  • Own the day-to-day partner relationship, building trust and resolving issues while protecting Transluce's core interests and the technical team's time.
  • Engage directly with evaluation results and technical analysis to understand the work, identify issues, and help drive the engagement forward.
  • Present deliverables to partners and summarize results for broader audiences, including blog posts, papers, and government and media briefings.

Minimum qualifications:
  • Understanding of AI evaluations and the ability to engage with technical teams on their design and results.
  • Track record of managing cross-functional teams, for example in a tech company, consulting firm, or fast-paced startup.
  • Experience engaging external partners such as frontier AI companies, government actors, or civil society organizations.
  • Ability to work through ambiguity and deliver against complex, evolving goals on short timelines, building workflows where none exist.
  • Ability to translate technical research for a range of audiences.
  • Exceptional communication and relationship skills, low ego, openness to giving and receiving feedback.

Preferred qualifications (not required):
  • Experience managing evaluations at scale or in a production context.
  • Proficiency in Python and data analysis tools (e.g., pandas, SQL), and familiarity with common ML and evaluation frameworks (e.g., PyTorch, Hugging Face, Inspect).
  • Experience building and applying software tools, for example in a forward-deployed engineering role.
  • Knowledge of frontier AI risk management practices or relevant regulatory guidance.

We are hiring at all levels of experience and would encourage those enthusiastic about the role who do not meet all of the qualifications to apply. This role involves regular travel. We are located in San Francisco and excited to work together in person. We are open to sponsoring international visas.

Skills Required

  • Understanding of AI evaluations and ability to engage technical teams on evaluation design and results
  • Track record managing cross-functional teams in a technology company, consulting firm, or fast-paced startup
  • Experience engaging external partners such as frontier AI companies, government actors, or civil society organizations
  • Ability to work through ambiguity and deliver complex, evolving goals on short timelines
  • Ability to build workflows where none exist
  • Ability to translate technical research for a range of audiences
  • Exceptional communication and relationship skills
  • Low ego and openness to giving and receiving feedback
  • Experience managing evaluations at scale or in a production context
  • Proficiency in Python and data analysis tools such as pandas and SQL
  • Familiarity with ML and evaluation frameworks such as PyTorch, Hugging Face, and Inspect
  • Experience building and applying software tools, such as in a forward-deployed engineering role
  • Knowledge of frontier AI risk management practices or relevant regulatory guidance
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: San Francisco, California
20 Employees
Year Founded: 2024

What We Do

Transluce is an independent research lab that builds open, scalable technology for understanding AI systems and steering them in the public interest. Transluce means to shine light through something to reveal its structure. Today’s complex AI systems are difficult to understand—not even experts can reliably predict their behavior once deployed. Given AI's extraordinary consequences on society, we need scalable and open analyses of the capabilities and risks of AI systems. We are building open source, AI-driven tools to understand and analyze AI systems. We will apply these tools to open-weight models, so the world can vet our analyses and improve their reliability. Once our technology has been vetted, we will work with frontier AI labs and governments to ensure that internal assessments reach the same standards as our publicly vetted procedures. Email: [email protected]

Similar Jobs

Hybrid
Irvine, CA, USA
289097 Employees

Snap Inc. Logo Snap Inc.

Design Engineer

Artificial Intelligence • Cloud • Machine Learning • Mobile • Software • Virtual Reality • App development
Hybrid
Palo Alto, CA, USA
5000 Employees
133K-235K Annually

Wipfli Logo Wipfli

Manager, Financial Reporting - Nonprofit Clients

Cloud • Fintech • Software • Business Intelligence • Consulting • Financial Services
Remote or Hybrid
United States
2900 Employees
97K-145K Annually

Wipfli Logo Wipfli

Manager, Accounting Advisory Services - Nonprofit Industry Clients

Cloud • Fintech • Software • Business Intelligence • Consulting • Financial Services
Remote or Hybrid
United States
2900 Employees
107K-160K Annually

Similar Companies Hiring

Revel Thumbnail
Aerospace • Hardware • Robotics • Software
Marina Del Rey, California
60 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software
New York, New York
30 Employees
Vega Thumbnail
Artificial Intelligence • Automotive • Insurance • Transportation
US
43 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account