Research Engineer

Posted 24 Days Ago
San Francisco, CA, USA
In-Office
180K-280K Annually
Entry level
Artificial Intelligence • Blockchain • Software • Cryptocurrency
The Role
Develops AI agents and language-model systems for automated code review and validation. Responsibilities include designing multi-agent architectures, building models, harnesses, and evaluations, researching LLMs and information retrieval, prototyping code-review workflows, and integrating successful experiments into production. The role requires strong programming, research ability, and product intuition, with opportunities to work on autonomous testing and deployment systems.
Summary Generated by Built In
Research Engineer

Greptile is an AI code reviewer that catches bugs and anti-patterns in pull requests with complete context of the codebase. Hundreds of top software companies, from YC startups to big tech, and teams in finance, healthcare, and defense, use Greptile to merge PRs faster and catch more bugs.

Greptile reviews 5B lines of code every month for 22,000+ customers. A new repo starts using Greptile every 2 minutes. We also uniquely offer self-hosted, air-gapped deployments for enterprises across defense, financial services, and healthcare.

We care more about hiring great people than filling specific slots, and everyone here wears multiple hats. The day-to-day for any given person tends to span more than one area.

Problems we’re excited about
  • Coding standards can be idiosyncratic and are often poorly documented; can we build agents that learn them through osmosis like a new hire might?

  • Can we identify for each customer what types of PR feedback they do and don’t care about, perhaps using some sample efficient RL, in order to increase signal-to-noise ratio?

  • Some bugs are best caught by running the code, potentially against discerning AI-generated E2E tests. Can we autonomously deploy feature branches and use agents to parallel try to break the application to detect bugs?

Trajectory
  • 22,000+ customers

  • 5B lines of code reviewed every month

  • Scaled from $0 to eight figures in ARR

  • Raised $30M from Benchmark, Y Combinator, Paul Graham, and Initialized

Team
  • We have assembled a small, talent dense team who have scaled critical functions at companies like Stripe, Google, Figma, etc.

Responsibilities
  • Experiment with and apply the latest advances in agents and language models to push the performance and capabilities of our products

  • Design and build the multi-agent systems, harnesses, models, and evals that power code validation

  • Example: you might study multi-agent architectures, prototype and evaluate a multi-agent code review workflow, and then work with a team to integrate successful prototypes into production systems

  • Stay on top of the latest research across LLMs, information retrieval, and developer tooling

Qualifications
  • B.S. in Computer Science or equivalent

  • Research experience in math, computer science, or physics

  • Strong programming skills and sharp product intuition

  • Bonus: research experience with ML, language models, or agents; papers accepted at top ML conferences; your own harnesses and evals for technical applications; experience post-training models; or a research master’s or PhD

You’ll like working here if
  • You want to work on a product that thousands of developers rely on

  • The chaos of high growth and things breaking is exciting to you

  • You like being in an office every day around other smart people building hard things

  • You love solving hard problems and shipping things that real users feel

  • You’re AI-enthusiastic, you hack on side projects, you write a blog (technical or not), and you’re excited about what you’re building.

Skills Required

  • B.S. in Computer Science or equivalent
  • Research experience in mathematics, computer science, or physics
  • Strong programming skills
  • Sharp product intuition
  • Research experience with machine learning, language models, or agents
  • Papers accepted at top machine learning conferences
  • Experience building harnesses and evaluations for technical applications
  • Experience post-training models
  • Research master's degree or Ph.D.
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: San Francisco, California
17 Employees
Year Founded: 2021

What We Do

AI expert that understands your codebase, as an API. Greptile can review your pull requests, answer questions about your codebase, write descriptions for JIRA tickets and more.

Similar Jobs

Hybrid
San Francisco, CA, USA
205000 Employees
87K-168K Annually

ServiceNow Logo ServiceNow

Scientist

Artificial Intelligence • Cloud • HR Tech • Information Technology • Productivity • Software • Automation
Hybrid
Santa Clara, CA, USA
29000 Employees
203K-354K Annually

ServiceNow Logo ServiceNow

Scientist

Artificial Intelligence • Cloud • HR Tech • Information Technology • Productivity • Software • Automation
Hybrid
Santa Clara, CA, USA
29000 Employees
232K-405K Annually

Benchling Logo Benchling

Research Engineer, Model Evaluation and Improvement

Cloud • Healthtech • Social Impact • Software • Biotech
Hybrid
San Francisco, CA, USA
605 Employees
136K-265K Annually

Similar Companies Hiring

Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees
Revel.io Thumbnail
Aerospace • Hardware • Robotics • Software
US
50 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software • Productivity
US
15 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account