×
Register Here to Apply for Jobs or Post Jobs. X

GPT Infra Technical Program Manager

Job in San Francisco, San Francisco County, California, 94199, USA
Listing for: Slope
Full Time position
Listed on 2026-08-28
Job specializations:
  • IT/Tech
  • Business
Salary/Wage Range or Industry Benchmark: 180000 - 240000 USD Yearly USD 180000.00 240000.00 YEAR
Job Description & How to Apply Below

About the Team

OpenAI’s Industrial Compute team is responsible for building and scaling the external infrastructure ecosystem that powers advanced AI systems. We work across hyperscalers, colocation providers, cloud partners, and strategic third-party operators to turn contracted capacity into production-ready compute.

Our scope spans the full lifecycle of external deployments: commercial alignment, technical readiness, network integration, hardware enablement, operational readiness, and long-range scaling strategy.

As OpenAI’s infrastructure footprint expands globally, we need leaders who can convert complex partner environments into reliable, high-velocity capacity for training and inference workloads.

About the Role

We are seeking a Technical Program Manager for our GPT Infrastructure teams, to lead delivery of external compute capacity that directly serves OpenAI model workloads.

In this role, you will own complex cross-functional programs that transform third-party infrastructure into usable tokens  will partner across engineering, capacity planning, networking, hardware, finance, product, and external providers to ensure that deployed capacity translates into real production throughput.

This role sits at the intersection of infrastructure execution, systems readiness, and business impact. Success requires strong technical fluency, elite program management, and the ability to drive accountability across internal teams and external partners.

This is a high-visibility role with direct impact on OpenAI’s ability to scale model training and inference globally.

This role is based in San Francisco, CA, with a hybrid work model of 3 days in office per week. Relocation assistance is available.

Key Responsibilities
  • Lead end-to-end delivery programs that convert external infrastructure capacity into production-ready token supply.
  • Own readiness across compute, storage, networking, security, and operational dependencies for third-party environments.
  • Build integrated plans across internal engineering teams and external partners with clear milestones, owners, risks, and critical paths.
  • Drive launch execution for new partner regions, clusters, and capacity expansions.
  • Create operating mechanisms that measure deployed capacity versus usable token output.
  • Identify bottlenecks preventing token generation (network constraints, hardware readiness, software enablement, partner delays, etc.) and drive resolution.
  • Coordinate with capacity planning and finance teams to prioritize the highest ROI capacity opportunities.
  • Establish executive-level reporting on delivery status, risks, and token ramp forecasts.
  • Improve repeatability of partner onboarding, technical integration, and scaling motions.
  • Manage escalations across internal and external stakeholders during high-severity delivery issues.
  • Translate ambiguous infrastructure constraints into clear execution plans.
  • Help define the long-term operating model for Token-as-a-Service across Stargate and 3P ecosystems.
Qualifications
  • 8+ years of Technical Program Management, Engineering Program Management, or Infrastructure Delivery experience.
  • Experience leading large-scale technical programs involving cloud, data center, networking, hardware, or distributed systems.
  • Strong understanding of compute infrastructure, clusters, networking, storage, and production systems.
  • Proven ability to drive cross-functional execution across engineering, operations, finance, and external vendors.
  • Experience managing executive stakeholders and communicating complex tradeoffs clearly.
  • Strong analytical skills with ability to reason about utilization, throughput, capacity, and operational metrics.
  • Comfortable operating in ambiguous, fast-scaling environments.
  • Strong written and verbal communication skills.
  • High ownership mentality with bias toward action.
  • Experience working with external providers, strategic partners, or hyperscalers is highly preferred.
Preferred Skills
  • Experience with GPU clusters, AI infrastructure, or large-scale model serving environments.
  • Familiarity with token economics, inference capacity planning, or workload scheduling.
  • Experience scaling global infrastructure through third-party providers.
  • Ba…
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary