Infrastructure Capacity Engineer

Reposted 5 Days Ago
Be an Early Applicant
2 Locations
In-Office
225K-300K
Mid level
Artificial Intelligence • Software
The Role
The Infrastructure Capacity Engineer will design capacity planning models, optimize resource utilization, and manage infrastructure for AI/ML workloads to support growth.
Summary Generated by Built In

Perplexity is an AI-powered answer engine founded in December 2022 and growing rapidly as one of the world’s leading AI platforms. Perplexity has raised over $1B in venture investment from some of the world’s most visionary and successful leaders, including Elad Gil, Daniel Gross, Jeff Bezos, Accel, IVP, NEA, NVIDIA, Samsung, and many more. Our objective is to build accurate, trustworthy AI that powers decision-making for people and assistive AI wherever decisions are being made. Throughout human history, change and innovation have always been driven by curious people. Today, curious people use Perplexity to answer more than 780 million queries every month–a number that’s growing rapidly for one simple reason: everyone can be curious. 

Perplexity is seeking an experienced Infrastructure Capacity Engineer to own our infrastructure scaling, capacity planning, and resource optimization across our AI/ML infrastructure. The ideal candidate will have deep experience in large-scale distributed systems, capacity modeling, and infrastructure efficiency optimization to support our rapidly growing AI products and user base.

Responsibilities
  • Design and implement comprehensive capacity planning models and forecasting systems that predict infrastructure needs across compute, storage, and network resources for our AI/ML workloads
  • Build and maintain automated capacity management systems that dynamically scale our infrastructure based on real-time demand patterns and usage forecasts
  • Lead cross-functional capacity planning initiatives including hardware procurement, data center expansion, and cloud resource optimization
  • Develop sophisticated monitoring and alerting systems that provide early warning indicators for capacity constraints and performance degradation
  • Create and maintain detailed infrastructure capacity models that account for seasonal patterns, product launches, and scaling efficiency across different workload types
  • Optimize resource utilization and cost efficiency through advanced placement algorithms, load balancing strategies, and infrastructure rightsizing
  • Design and implement disaster recovery and business continuity plans that ensure service availability during infrastructure failures or capacity emergencies
  • Collaborate with Site Reliability Engineering and Platform teams to establish capacity-aware deployment strategies and infrastructure automation
  • Play a leading role in defining the capacity engineering discipline within Perplexity’s engineering organization
Qualifications
  • Minimum of 4+ years of experience in infrastructure capacity planning, systems engineering, or related technical roles at scale
  • Proven experience managing infrastructure capacity for high-growth technology companies, preferably with AI/ML workloads or real-time systems
  • Strong background in distributed systems architecture, cloud infrastructure (AWS/GCP/Azure), and container orchestration (Kubernetes)
  • Experience with capacity modeling tools, forecasting methodologies, and statistical analysis for infrastructure planning
  • Proficiency in programming languages such as Python, Go, or similar for automation and tooling development
  • Deep understanding of infrastructure monitoring, observability, and performance optimization techniques
  • Experience with infrastructure-as-code tools (Terraform, Ansible) and CI/CD pipelines for infrastructure management
  • Strong analytical and problem-solving skills with the ability to make data-driven decisions under uncertainty
  • Excellent cross-functional collaboration skills and experience working with engineering, product, and business stakeholders
  • Experience with large-scale database systems, caching layers, and content delivery networks preferred
  • Background in AI/ML infrastructure, LLM inference, GPU cluster management, or high-performance computing is a plus

Our cash compensation range for this role is $225,000 - $300,000.


Final offer amounts are determined by multiple factors, including, experience and expertise, and may vary from the amounts listed above.
 
Equity: In addition to the base salary, equity may be part of the total compensation package.
Benefits: Comprehensive health, dental, and vision insurance for you and your dependents. Includes a 401(k) plan.
 
 

Top Skills

Ansible
AWS
Azure
GCP
Go
Kubernetes
Python
Terraform
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: San Francisco, California
41 Employees
Year Founded: 2022

What We Do

The most powerful answer engine. Powering curiosity with answers backed by up-to-date sources. This is where knowledge begins.

Similar Jobs

In-Office
Pleasanton, CA, USA
14894 Employees
184K-327K Annually

Circle Logo Circle

Senior Site Reliability Engineer

Blockchain • Fintech • Payments • Financial Services • Cryptocurrency • Web3
In-Office
San Diego, CA, USA
980 Employees
148K-195K Annually

Circle Logo Circle

Senior Site Reliability Engineer

Blockchain • Fintech • Payments • Financial Services • Cryptocurrency • Web3
In-Office
Los Angeles, CA, USA
980 Employees
148K-195K Annually

Circle Logo Circle

Senior Site Reliability Engineer

Blockchain • Fintech • Payments • Financial Services • Cryptocurrency • Web3
In-Office
San Francisco, CA, USA
980 Employees
148K-195K Annually

Similar Companies Hiring

Standard Template Labs Thumbnail
Software • Information Technology • Artificial Intelligence
New York, NY
10 Employees
PRIMA Thumbnail
Travel • Software • Marketing Tech • Hospitality • eCommerce
US
15 Employees
Scotch Thumbnail
Software • Retail • Payments • Fintech • eCommerce • Artificial Intelligence • Analytics
US
25 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account