ForgeApply · Job listing
Hardware Technical Program Manager, Infrastructure Partner Operations
Openai
Apply in about a minute — without sacrificing quality.
ForgeApply autofills this application and tailors your resume to this exact posting. You review everything before it's sent. Free trial, no card required.
About this role
ABOUT THE TEAM
The Industrial Compute team is responsible for building the physical infrastructure that powers OpenAI’s largest-scale AI systems. We design, deploy, and operate next-generation compute infrastructure across a rapidly expanding global footprint, combining OpenAI-owned infrastructure with strategic cloud and infrastructure partners to support frontier AI workloads.
As our infrastructure footprint grows, operational excellence across third-party providers becomes increasingly critical. Our team ensures external infrastructure partners consistently deliver the reliability, performance, and operational maturity required to support OpenAI’s rapidly expanding compute environment.
About the Role
We are seeking a Hardware Technical Program Manager, Infrastructure Partner Operations to lead operational delivery across OpenAI’s third-party infrastructure partners, including major cloud service providers and strategic compute vendors.
In this role, you will serve as the primary operational program manager for external infrastructure partners, driving accountability for service delivery, operational readiness, incident management, performance reporting, and continuous operational improvement. You will work closely with partner engineering and operations teams while coordinating internally across Hardware Engineering, Infrastructure Operations, Capacity Planning, Networking, Supply Chain, Deployment, Reliability Engineering, and executive leadership.
Success in this role requires someone who understands how hyperscale infrastructure organizations operate, can establish strong operational governance with external partners, and is comfortable driving complex technical programs without direct ownership of the underlying infrastructure.
Key Responsibilities
- Own operational engagement with third-party infrastructure providers, ensuring consistent execution against operational commitments, service-level agreements (SLAs), and performance expectations.
- Develop operational governance frameworks with strategic partners, including business reviews, operational scorecards, escalation processes, executive reporting, and performance improvement plans.
- Define, track, and continuously improve key operational metrics related to infrastructure availability, deployment execution, incident response, operational health, service quality, and partner performance.
- Build dashboards and reporting mechanisms that provide clear visibility into partner operational performance, risks, trends, and areas requiring executive attention.
- Drive cross-functional coordination between OpenAI teams and external infrastructure providers to resolve operational issues, remove execution blockers, and improve delivery outcomes.
- Lead operational escalations involving infrastructure availability, deployment execution, hardware operations, capacity delivery, or service performance, ensuring timely resolution and clear executive communication.
- Establish repeatable operating rhythms with external partners, including weekly operational reviews, executive business reviews, service reviews, action tracking, and long-term improvement initiatives.
- Partner with Capacity Planning, Hardware Operations, Networking, Deployment, Reliability Engineering, and Supply Chain teams to ensure external infrastructure providers remain aligned with OpenAI’s operational priorities.
- Identify systemic operational risks across partner organizations and proactively drive corrective actions that improve long-term operational effectiveness.
Qualifications
- 7+ years of experience in Technical Program Management, Infrastructure Operations, Cloud Operations, Service Delivery, or Technical Account Management within large-scale infrastructure environments.
- Experience managing operational relationships with external infrastructure providers, cloud service providers, hardware vendors, or strategic technology partners.
- Strong understanding of hyperscale cloud infrastructure, data center operations, infrastructure delivery, or large-scale distributed systems.
- Experience developing operational KPIs, SLAs, service health metrics, dashboards, and executive reporting for complex technical organizations.
- Demonstrated success leading cross-functional operational programs involving both internal stakeholders and external partners.
- Strong program management skills with the ability to drive accountability across organizations without direct authority.
- Excellent written and verbal communication skills with experience presenting operational performance to senior technical and executive leadership.
- Bachelor's degree in Engineering, Computer Science, Information Systems, Operations, or equivalent practical experience.
Preferred Skills
- Experience managing cloud infrastructure operations within organizations such as Microsoft Azure, Amazon Web Services (AWS), Google Cloud Platform (GCP), Oracle Cloud Infrastructure (OCI), or other hyperscale cloud providers.
- Experience leading operational governance, service delivery, customer success engineering, technical account management, or infrastructure operations for enterprise cloud customers.
- Strong understanding of service-level agreements (SLAs), operational KPIs, incident management, escalation processes, root cause analysis, and continuous service improvement methodologies.
- Experience building executive dashboards, operational scorecards, business review frameworks, and data-driven performance reporting.
- Familiarity with infrastructure operations supporting GPU infrastructure, AI infrastructure, high-performance computing (HPC), or hyperscale data center environments.
- Experience managing complex cross-company technical relationships while balancing customer priorities, engineering constraints, and operational execution.
- Proven ability to influence senior stakeholders across both internal teams and external partner organizations withou
Salary insight
The midpoint of this range ($256k) is about 26% above the median disclosed salary for San Francisco roles listed on ForgeApply ($203k across 6,275 jobs).
See full Project Manager salary data for San Francisco →
Based on live postings with disclosed pay on ForgeApply; refreshed daily. Not an estimate of this employer's offer.
Ready to apply to Openai?
Apply in about a minuteSimilar jobs
- Technical Program Manager, Infrastructure — Etched · San Jose
- Technical Program Manager, Infrastructure — Figma · San Francisco, CA • New York, NY • United States
- Technical Program Manager, Infrastructure — Baseten · San Francisco
- Technical Program Manager, Infrastructure — Anthropic · San Francisco, CA | New York City, NY | Seattle, WA
- Senior Technical Program Manager, Infrastructure — Lyft · San Francisco, CA
- Technical Program Manager, Service Infrastructure — Stripe · Remote
- Technical Program Manager, Infra — Oscar · New York, New York, United States
- Technical Program Manager (TPM), Infrastructure — Cursor · San Francisco
More like this: Project & Program Manager Jobs · Project & Program Manager Jobs in San Francisco · More jobs at Openai · Browse all jobs