jobloom

JobLoom finds jobs directly from company career sites before many job boards, then routes you into detailed role pages like this one.

other

Posted Apr 21

Token-as-a-Service Technical Program Manager

at openai

San Francisco, United StatesHybrid

Responsibilities

  • - Own readiness across compute, storage, networking, security, and operational dependencies for third-party environments.
  • - Build integrated plans across internal engineering teams and external partners with clear milestones, owners, risks, and critical paths.
  • - Drive launch execution for new partner regions, clusters, and capacity expansions.
  • - Create operating mechanisms that measure deployed capacity versus usable token output.

Requirements

  • About the Team OpenAI’s Stargate and 3P Engineering teams are responsible for building and scaling the external infrastructure ecosystem that powers advanced AI systems.
  • As OpenAI’s infrastructure footprint expands globally, we need leaders who can convert complex partner environments into reliable, high-velocity capacity for training and inference workloads.
  • About the Role We are seeking a Technical Program Manager, Token-as-a-Service (TaaS) to lead delivery of external compute capacity that directly serves OpenAI model workloads.
  • Success requires strong technical fluency, elite program management, and the ability to drive accountability across internal teams and external partners.
  • This is a high-visibility role with direct impact on OpenAI’s ability to scale model training and inference globally.
  • Experience leading large-scale technical programs involving cloud, data center, networking, hardware, or distributed systems. - Strong understanding of compute infrastructure, clusters, networking, storage, and production systems. - Proven ability to drive cross-functional execution across engineering, operations, finance, and external vendors. -
  • Experience managing executive stakeholders and communicating complex tradeoffs clearly. - Strong analytical skills with ability to reason about utilization, throughput, capacity, and operational metrics. - Comfortable operating in ambiguous, fast-scaling environments. - Strong written and verbal communication skills. - High ownership mentality with bias toward action. -
  • Experience working with external providers, strategic partners, or hyperscalers is highly preferred. Preferred Skills -
  • Experience with GPU clusters, AI infrastructure, or large-scale model serving environments. - Familiarity with token economics, inference capacity planning, or workload scheduling. -
  • Experience building new operational models in high-growth environments. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence

Experience

  • Qualifications - 8+ years of Technical Program Management, Engineering Program Management, or Infrastructure Delivery experience. -

Additional details

  • We work across hyperscalers, colocation providers, cloud partners, and strategic third-party operators to turn contracted capacity into production-ready compute.
  • Our scope spans the full lifecycle of external deployments: commercial alignment, technical readiness, network integration, hardware enablement, operational readiness, and long-range scaling strategy.
  • In this role, you will own complex cross-functional programs that transform third-party infrastructure into usable tokens at scale.
  • You will partner across engineering, capacity planning, networking, hardware, finance, product, and external providers to ensure that deployed capacity translates into real production throughput.
  • This role sits at the intersection of infrastructure execution, systems readiness, and business impact.
  • This role is based in San Francisco, CA, with a hybrid work model of 3 days in office per week.
  • Responsibilities - Lead end-to-end delivery programs that convert external infrastructure capacity into production-ready token supply.
  • - Establish executive-level reporting on delivery status, risks, and token ramp forecasts.
  • - Improve repeatability of partner onboarding, technical integration, and scaling motions.
  • - Manage escalations across internal and external stakeholders during high-severity delivery issues.

Find more real-time jobs on JobLoom.