June 2026 archive · For research and skill exploration. Current job availability is unverified.
OpenAI
Technical Program Manager, Compute Infrastructure
Lead end-to-end GPU cluster delivery across partners, managing hardware, networking, and capacity at scale.
OpenAIText summary from the June 2026 archive. original source →
The role
This Technical Program Manager role oversees large-scale GPU cluster deployment and compute infrastructure that powers ChatGPT and training workloads. Reporting into an engineer-first TPM team, the role bridges hardware providers, engineering teams, and leadership to bring production-ready capacity online reliably and at massive scale.
What you'd do
- Own end-to-end delivery of new compute SKUs and GPU clusters across external partner networks
- Drive multi-threaded bring-up programs spanning hardware, networking, power, and cooling infrastructure
- Interface with chip vendors to derisk onboarding of new hardware platforms
- Build and operationalize program mechanisms including roadmaps, milestones, and risk registers
- Partner with engineers to improve cluster reliability, repeatability, and automation
- Support physical and logical bring-up of network Points-of-Presence including rack deployment
- Coordinate cross-functional readiness across security, finance, operations, and research teams
- Manage handoffs and integration across teams and external partners
- Identify bottlenecks and drive systemic process improvements
- Provide executive visibility on portfolio progress, tradeoffs, and risks
What they're looking for
- Degree in hard science or demonstrated engineering expertise
- 5+ years program management experience on major capital or hyperscaler infrastructure projects
- Proven ability to own and deliver complex projects independently
- Experience managing cross-functional and cross-company teams with strong communication discipline
- Track record delivering high-profile technical projects under tight deadlines
- Technical aptitude and history partnering with top-tier engineering or research teams
- Experience leading external vendors including engineering firms and equipment suppliers
- Expertise designing and implementing scalable processes for complex problems
- Experience managing complicated dependencies in logistics and supply chains
- Resourcefulness and comfort thriving in ambiguous, fast-paced environments
Nice to have
- Thoughtful perspective on impacts of artificial general intelligence
Summary written by RoleDeck from the original posting. This is an extracted, own-words summary and may contain errors. The original source may have changed or expired.