CM
The Career Machine
Back to Jobs
Ashby Verified ATS

Technical Program Manager, Model Deployment & Capacity

Openai San Francisco 2w ago
FullTime
Analyzing job security signals...
Employer Website Verification
jobs.ashbyhq.com

About Openai (Wikipedia Record):

OpenAI is an American artificial intelligence (AI) public benefit corporation (PBC) headquartered in San Francisco, California. It develops proprietary generative AI models, particularly its generative pre-trained transformer (GPT) series of large language models. Its release of ChatGPT in November 2022 has been credited with catalyzing the AI boom; as of September 2026, ChatGPT is the fifth-most-visited website globally. OpenAI also releases the GPT Image models and Codex, an AI coding agent. In March 2026, OpenAI closed a funding round with a post-money valuation of US$852 billion, making it one of the most valuable AI pure play companies in the world, rivalled only by Anthropic.

About the Role

ABOUT THE TEAM The Product & Platform teams at OpenAI are responsible for delivering the company’s most impactful offerings—such as ChatGPT, our API platform, and new enterprise capabilities—to a global and diverse customer base. These systems must perform at scale and deliver exceptional experiences to developers, consumers, and businesses alike. The ChatGPT infrastructure team is responsible for ensuring that our products can serve rapidly growing demand with the performance, reliability, and quality our users expect. This work sits at the intersection of product demand, model deployment, inference, research, fleet, and capacity. The team translates changing product and model needs into clear capacity decisions and safe, scalable launches.   ABOUT THE ROLE We are seeking a Technical Program Manager to lead the operating system for Chat capacity and model deployment. You will connect demand forecasting and capacity allocation with model readiness, rollout planning, launch coordination, and post-deployment learning. You will also own mode deployment beyond capacity by working with cross functional teams across research, post-training, inference and product to own mainline model deployment. You will bring structure to constrained-capacity decisions, improve the tooling and mechanisms teams use to prioritize demand, and help new models reach users safely and efficiently. Success requires technical depth, sound judgment under ambiguity, and crisp execution across product, research, infrastructure, and operations teams. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees.   IN THIS ROLE, YOU WILL: - Own cross-functional programs for Chat capacity forecasting, allocation, headroom planning, and constrained-capacity operations. - Build durable intake, prioritization, and decision mechanisms that connect product demand and model requirements to available servin

Similar Roles

View all
C
CloudflareHybrid
ExcelProblem SolvingLeadershipCommunication
Remote