Technical Program Manager
Runpod · Remote - USA · United States · Remote
Pay: USD 140,000 – 165,000 a year
Posted Aug 26, 2026
Sign up free: we match you to jobs like this, tailor your application and fill the form. 2 free applications every day.
Runpod is the AI Developer Cloud. More than one million developers, from indie researchers to teams running frontier models in production, use Runpod to experiment, train, fine-tune, deploy, and scale AI on one platform. The platform has processed more than 20 billion inference requests. We closed a $100M Series A in June 2026. We're at an inflection point for AI infrastructure, and we're building the platform the next generation of developers will depend on.
We're a small, remote-first team. We take ownership seriously, move fast, and ship work that more than a million developers rely on every day. We're looking for people who care deeply, build with urgency, and want to matter at scale.
Learn more in our CEO's funding announcement: https://www.runpod.io/blog/one-million-developers.
We're seeking a Technical Program Manager (TPM) who combines strong technical depth with exceptional organizational and communication skills. The ideal candidate is a self-starter who thrives in a fast-paced, high-growth environment and drives execution through ambiguity with confidence and clarity.
As a TPM at RunPod, you'll play a central role in leading complex, cross-functional programs across product and platform teams. This role sits inside our Supply organization — you'll be the connective tissue between Engineering, Product, Data Center Management, and our host partner teams, working to make sure GPU capacity keeps pace with demand. Through strong communication, alignment, and proactive risk management, you'll drive the successful delivery of high-impact programs that keep supply, cost, and reliability in balance for our customers.
Responsibilities:
1. Partner with Data Center Management and Host/Supply teams to run the capacity and utilization review cadence that feeds Supply's input into the product roadmap — surfacing constraints early (host churn, new capacity coming online, regional gaps) and turning them into actionable planning input.
2. Own the relationship and escalation path with host and infrastructure partners — resolving capacity, performance, or contractual issues quickly, and keeping Engineering and Product informed of anything that could affect delivery.Own the relationship and escalation path with host and infrastructure partners — resolving capacity, performance, or contractual issues quickly, and keeping Engineering and Product informed of anything that could affect delivery.
3. Translate capacity and utilization data into concrete tradeoffs: where to add supply, where to consolidate, and how each option nets out on cost and reliability.
4. Build and drive project plans across Engineering, Product, and Supply stakeholders — scoping work, sequencing dependencies, and setting timelines that hold up against real-world constraints like partner lead times and hardware delivery.
5. Act as the primary point of contact across these teams: spot risks and roadblocks early, and keep stakeholders aligned with clear, regular…