Site Reliability / Infrastructure Engineer

Posted Yesterday
Be an Early Applicant
New York City, NY, USA
In-Office
180K-275K Annually
Mid level
Artificial Intelligence • Machine Learning • Robotics • Generative AI
The Role
Own on-call and incident response for large-scale video ingestion and social infrastructure. Drive reliability, scaling (DB sharding, Elasticsearch), IaC with Terraform, GCP/Kubernetes operations, CI/CD, and lead postmortems and cross-team infra work.
Summary Generated by Built In
About the Company

General Intuition is the frontier lab for acting in space and time. We build large action models and world models that can perceive, predict, and act across virtual and physical environments. General Intuition builds on the strength of Medal, the world's largest and fastest-growing platform for gaming clips, where millions of gamers capture, share, and discover new games every year. We recently raised $320M at a $2.3B valuation led by Khosla Ventures with participation from General Catalyst, Eric Schmidt, and Jeff Bezos, to discover the next generation of real-world intelligence.

The Role

Medal's infrastructure handles billions of clips, video ingestion pipelines, and social features at a massive scale most engineers never get to touch. The work centers on reliability, incident response, scaling, and making sure our infrastructure keeps up with our growth. You'll own the on-call rotation, drive postmortems, and work directly with engineering teams to meet their infra needs. The right person probably came through startups and scale-ups, has been in the room when things broke at 2am, has scaled databases under pressure, and knows the difference between a durable fix and a patch that buys you a week.

What We're Looking For
  • Infrastructure-as-code: Strong fluency in Terraform, with real experience owning infrastructure-as-code at scale

  • Elasticsearch depth: Hands-on experience running ES for user-facing features, not just as a log sink

  • GCP depth: You know it maybe a little too well: Kubernetes, VPC, IAM, Cloud Logging, and the managed services ecosystem

  • Database scaling: Deep, hands-on experience scaling and sharding relational databases (MySQL, Postgres) in production

  • Incident response instincts: You can work a P0 calmly, communicate clearly under pressure, and run a postmortem that prevents recurrence

  • CI/CD: You've worked with GitHub Actions in a production environment

  • Communication (crucial!): You flag issues clearly and rapidly during incidents and lead/write actionable postmortems

  • Experience at startups: You are comfortable in an environment of rapid growth where scaling up is a priority

  • Great judgment: You know the difference between a durable, sustainable fix and a patch that buys you a week

Our Stack

Electron, React, Redux, Styled Components & other modern web-based technologies
C# and C++ for native Windows recording & more
Swift for iOS, Kotlin for Android
Java, Redis, RabbitMQ, Kubernetes for backend
Terraform, Salt, GitHub Actions, CircleCI for IaC and CI/CD

Skills Required

  • Strong fluency in Terraform and owning infrastructure-as-code at scale
  • Hands-on experience running Elasticsearch for user-facing features
  • Deep GCP knowledge (Kubernetes, VPC, IAM, Cloud Logging, managed services)
  • Deep, hands-on experience scaling and sharding relational databases (MySQL, Postgres)
  • Strong incident response skills, able to handle P0s and run postmortems
  • Experience with CI/CD in production, specifically GitHub Actions
  • Excellent communication for incident escalation and postmortem leadership
  • Experience at startups or scale-ups and comfort with rapid growth environments
  • Sound judgment to choose durable fixes over temporary patches
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
41 Employees
Year Founded: 2025

What We Do

General Intuition is an AI research lab and public-benefit corporation building frontier and world models that help machines predict and act in dynamic environments. Using billions of action-labelled gameplay clips from its sister company Medal, the lab trains systems on physics, human decision-making, and consequences. Its mission extends to agentic models, robotics, and AI agents inhabiting real and imaginary worlds.

Similar Jobs

Treeswift Logo Treeswift

Infrastructure Engineer

Artificial Intelligence • Hardware • Machine Learning • Robotics • Software • Utilities
Hybrid
New York, NY, USA
55 Employees
160K-220K Annually

Oscar Health Logo Oscar Health

Senior Software Engineer

Healthtech • Insurance
In-Office
New York, NY, USA
2200 Employees
181K-237K Annually

Medal Logo Medal

Infrastructure Engineer

Gaming • Mobile • Software
In-Office
New York City, NY, USA
57 Employees
180K-275K Annually

MongoDB Logo MongoDB

Site Reliability Engineer

Big Data • Cloud • Software • Database
Easy Apply
Remote or Hybrid
5 Locations
5550 Employees
127K-249K Annually

Similar Companies Hiring

Hanover Park Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
42 Employees
LTX Thumbnail
Robotics • Conversational AI • Generative AI
Jerusalem, Israel
300 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account