other
Posted Apr 21Token-as-a-Service Technical Program Manager
at openai
San Francisco, United StatesHybrid
Responsibilities
- - Own readiness across compute, storage, networking, security, and operational dependencies for third-party environments.
- - Build integrated plans across internal engineering teams and external partners with clear milestones, owners, risks, and critical paths.
- - Drive launch execution for new partner regions, clusters, and capacity expansions.
- - Create operating mechanisms that measure deployed capacity versus usable token output.
Requirements
- About the Team OpenAI’s Stargate and 3P Engineering teams are responsible for building and scaling the external infrastructure ecosystem that powers advanced AI systems.
- As OpenAI’s infrastructure footprint expands globally, we need leaders who can convert complex partner environments into reliable, high-velocity capacity for training and inference workloads.
- About the Role We are seeking a Technical Program Manager, Token-as-a-Service (TaaS) to lead delivery of external compute capacity that directly serves OpenAI model workloads.
- Success requires strong technical fluency, elite program management, and the ability to drive accountability across internal teams and external partners.
- This is a high-visibility role with direct impact on OpenAI’s ability to scale model training and inference globally.
- Experience leading large-scale technical programs involving cloud, data center, networking, hardware, or distributed systems. - Strong understanding of compute infrastructure, clusters, networking, storage, and production systems. - Proven ability to drive cross-functional execution across engineering, operations, finance, and external vendors. -
- Experience managing executive stakeholders and communicating complex tradeoffs clearly. - Strong analytical skills with ability to reason about utilization, throughput, capacity, and operational metrics. - Comfortable operating in ambiguous, fast-scaling environments. - Strong written and verbal communication skills. - High ownership mentality with bias toward action. -
- Experience working with external providers, strategic partners, or hyperscalers is highly preferred. Preferred Skills -
- Experience with GPU clusters, AI infrastructure, or large-scale model serving environments. - Familiarity with token economics, inference capacity planning, or workload scheduling. -
- Experience building new operational models in high-growth environments. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence
Experience
- Qualifications - 8+ years of Technical Program Management, Engineering Program Management, or Infrastructure Delivery experience. -
Additional details
- We work across hyperscalers, colocation providers, cloud partners, and strategic third-party operators to turn contracted capacity into production-ready compute.
- Our scope spans the full lifecycle of external deployments: commercial alignment, technical readiness, network integration, hardware enablement, operational readiness, and long-range scaling strategy.
- In this role, you will own complex cross-functional programs that transform third-party infrastructure into usable tokens at scale.
- You will partner across engineering, capacity planning, networking, hardware, finance, product, and external providers to ensure that deployed capacity translates into real production throughput.
- This role sits at the intersection of infrastructure execution, systems readiness, and business impact.
- This role is based in San Francisco, CA, with a hybrid work model of 3 days in office per week.
- Responsibilities - Lead end-to-end delivery programs that convert external infrastructure capacity into production-ready token supply.
- - Establish executive-level reporting on delivery status, risks, and token ramp forecasts.
- - Improve repeatability of partner onboarding, technical integration, and scaling motions.
- - Manage escalations across internal and external stakeholders during high-severity delivery issues.