Runpod is the AI Developer Cloud. More than one million developers, from indie researchers to teams running frontier models in production, use Runpod to experiment, train, fine-tune, deploy, and scale AI on one platform. The platform has processed more than 20 billion inference requests. We closed a $100M Series A in June 2026. We're at an inflection point for AI infrastructure, and we're building the platform the next generation of developers will depend on.
We're a small, remote-first team. We take ownership seriously, move fast, and ship work that more than a million developers rely on every day. We're looking for people who care deeply, build with urgency, and want to matter at scale.
Learn more in our CEO's funding announcement: https://www.runpod.io/blog/one-million-developers.
We're seeking a Technical Program Manager (TPM) who combines strong technical depth with exceptional organizational and communication skills. The ideal candidate is a self-starter who thrives in a fast-paced, high-growth environment and drives execution through ambiguity with confidence and clarity.
As a TPM at RunPod, you'll play a central role in leading complex, cross-functional programs across product and platform teams. This role sits inside our Supply organization — you'll be the connective tissue between Engineering, Product, Data Center Management, and our host partner teams, working to make sure GPU capacity keeps pace with demand. Through strong communication, alignment, and proactive risk management, you'll drive the successful delivery of high-impact programs that keep supply, cost, and reliability in balance for our customers.
Responsibilities:
Partner with Data Center Management and Host/Supply teams to run the capacity and utilization review cadence that feeds Supply's input into the product roadmap — surfacing constraints early (host churn, new capacity coming online, regional gaps) and turning them into actionable planning input.
Own the relationship and escalation path with host and infrastructure partners — resolving capacity, performance, or contractual issues quickly, and keeping Engineering and Product informed of anything that could affect delivery.Own the relationship and escalation path with host and infrastructure partners — resolving capacity, performance, or contractual issues quickly, and keeping Engineering and Product informed of anything that could affect delivery.
Translate capacity and utilization data into concrete tradeoffs: where to add supply, where to consolidate, and how each option nets out on cost and reliability.
Build and drive project plans across Engineering, Product, and Supply stakeholders — scoping work, sequencing dependencies, and setting timelines that hold up against real-world constraints like partner lead times and hardware delivery.
Act as the primary point of contact across these teams: spot risks and roadblocks early, and keep stakeholders aligned with clear, regular status updates.
Bring technical judgment to evaluations of infrastructure tools, platforms, and vendor products, recommending what best fits RunPod's scale and constraints.
Run post-program reviews on completed initiatives, capturing what worked, what didn't, and what should change for the next cycle.
Requirements:
Previous experience in infrastructure-focused organizations, with exposure to capacity planning, data center operations, or vendor/partner management.
4+ years of experience or equivalent expertise in technical program management, leading complex technology programs from planning through delivery.
Proven track record of partnering with 3+ cross-functional teams to deliver high-impact programs from planning through launch.
Proven track record of working with external infrastructure or hardware vendors globally.
Strong technical background in engineering and scalable SaaS/IaaS products or services.
Strong quantitative and analytical skills, with the ability to turn capacity, cost, and delivery data into clear planning recommendations.
Deep understanding of the software development lifecycle (SDLC) and how engineering teams design, build, test, and ship products.
Highly organized and detail-oriented, with a track record of catching issues before they become blockers.
Demonstrated ability to independently build, track, and execute programs in dynamic environments with multiple dependencies, competing priorities, and tight deadlines.
A collaborative team player who values ownership, clarity, accountability, and follow-through over rigid process.
Exceptional verbal and written communication skills, with the ability to align stakeholders across teams and functions.
Solution-oriented and adaptable, comfortable navigating ambiguity and resolving challenges creatively under pressure.
Preferred:
Previous experience in AI/ML developer platforms, GPU/cloud infrastructure, or other hardware-capacity-constrained environments.
Experience at an early-stage or high-growth startup.
Technical degree in Computer Science, Engineering, or a related field.
What You’ll Receive:
The competitive base pay for this position ranges from ($140,000 - $165,000). This salary range may be inclusive of several career levels at Runpod and will be narrowed during the interview process based on a number of factors, including the candidate’s experience, qualifications, and location
Meaningful equity in a fast-growing company- everyone on the team receives stock options — your impact drives our growth, and you share in the upside.
Generous medical, dental & vision plans
Flexible PTO- take the time you need to recharge
Most roles are remote work first with an inclusive, collaborative teams utilizing slack as the main form of internal communication
Join a passionate team on the cutting edge of AI infrastructure — where culture, learning, and ownership are at the heart of how we scale.
• • $1,200 Home Office & Equipment Stipend- We set you up for success from day one with gear and support to create your ideal workspace
Runpod is committed to maintaining a workplace free from discrimination and upholding the principles of equality and respect for all individuals. We believe that diversity in all its forms enhances our team. As an equal opportunity employer, Runpod is committed to creating an inclusive workforce at every level. We evaluate qualified applicants without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, age, marital status, protected veteran status, disability status, or any other characteristic protected by law. We welcome every qualified candidate eligible to work in the United States; however, we are currently unable to sponsor employment visas.
Skills Required
- Previous experience in infrastructure-focused organizations, including capacity planning, data center operations, or vendor/partner management
- 4+ years of experience or equivalent expertise in technical program management leading complex technology programs
- Proven track record of partnering with 3+ cross-functional teams to deliver high-impact programs from planning through launch
- Proven track record of working with external infrastructure or hardware vendors globally
- Strong technical background in engineering and scalable SaaS/IaaS products or services
- Strong quantitative and analytical skills to turn capacity, cost, and delivery data into planning recommendations
- Deep understanding of the software development lifecycle (SDLC) and engineering workflows
- Highly organized and detail-oriented, with a track record of catching issues before they become blockers
- Demonstrated ability to independently build, track, and execute programs with multiple dependencies and tight deadlines
- Collaborative team player who values ownership, clarity, accountability, and follow-through
- Exceptional verbal and written communication skills to align stakeholders across teams
- Solution-oriented and adaptable, comfortable navigating ambiguity and resolving challenges under pressure
- Previous experience in AI/ML developer platforms, GPU/cloud infrastructure, or other hardware-capacity-constrained environments
- Experience at an early-stage or high-growth startup
- Technical degree in Computer Science, Engineering, or a related field
Runpod Compensation & Benefits Highlights
-
Healthcare Strength — Core medical, dental, and vision coverage is described, including fully paid health benefits for full-time employees. Feedback suggests the health plan is viewed favorably.
-
Retirement Support — A 401(k) with company matching is part of the package. Feedback suggests retirement support is a meaningful component of total rewards.
-
Equity Value & Accessibility — Stock options are offered broadly, making ownership accessible across roles. Feedback suggests equity is a notable element of total compensation.
Runpod Insights
What We Do
Runpod is pioneering the future of AI and machine learning, offering cutting-edge cloud infrastructure for full-stack AI applications. Founded in 2022, we are a rapidly growing, well-funded company with a remote-first organization spread globally. Our mission is to empower innovators and enterprises to unlock AI's true potential, driving technology and transforming industries. Join us as we shape the future of AI. We are building Cloud services focused on accelerating AI adoption. Whether you're an experienced ML developer training a large language model, or an enthusiast tinkering with stable diffusion, we strive to make GPU compute as seamless and affordable as possible.
Why Work With Us
Our Guiding Virtues Give a sh*t - We want to work with people who care - about our customers and about each other. Look in the mirror - We deeply reflect on our own actions and seek to better ourselves. Courage over comfort - We tackle hard truths and tough situations directly, even when it makes us uncomfortable.
Gallery
Runpod Offices
Remote Workspace
Employees work remotely.
We’re remote-first, offering flexibility with virtual tools for collaboration. For those nearby, we have coworking spaces in SF and Seattle. Enjoy the choice of office or remote work, with a focus on flexibility and work-life balance






