
Explore active software engineering, data, AI, product, and design roles directly from verified employers—and make sure your resume is ready before applying.
Get an instant ATS score, missing keyword alert, and bullet rewrites tailored to your target job before submitting your application.
About the Team OpenAI’s Industrial Compute team is responsible for building and scaling the external infrastructure ecosystem that powers advanced AI systems. We work across hyperscalers, colocation providers, cloud partners, and strategic third-party operators to turn contracted capacity into production-ready compute. Our scope spans the full lifecycle of external deployments: commercial alignment, technical readiness, network integration, hardware enablement, operational readiness, and long-range scaling strategy. As OpenAI’s infrastructure footprint expands globally, we need leaders who can convert complex partner environments into reliable, high-velocity capacity for training and inference workloads. About the Role We are seeking a Technical Program Manager for our GPT Infrastructure teams, to lead delivery of external compute capacity that directly serves OpenAI model workloads. In this role, you will own complex cross-functional programs that transform third-party infrastructure into usable tokens at scale. You will partner across engineering, capacity planning, networking, hardware, finance, product, and external providers to ensure that deployed capacity translates into real production throughput. This role sits at the intersection of infrastructure execution, systems readiness, and business impact. Success requires strong technical fluency, elite program management, and the ability to drive accountability across internal teams and external partners. This is a high-visibility role with direct impact on OpenAI’s ability to scale model training and inference globally. This role is based in San Francisco, CA, with a hybrid work model of 3 days in office per week. Relocation assistance is available. Key Responsibilities Lead end-to-end delivery programs that convert external infrastructure capacity into production-ready token supply. Own readiness across compute, storage, networking, security, and operational dependencies for third-party environments. Build integrated plans across internal engineering teams and external partners with clear milestones, owners, risks, and critical paths. Drive launch execution for new partner regions, clusters, and capacity expansions. Create operating mechanisms that measure deployed capacity versus usable token output. Identify bottlenecks preventing token generation (network constraints, hardware readiness, software enablement, partner delays, etc.) and drive resolution. Coordinate with capacity planning and finance teams to prioritize the highest ROI capacity opportunities. Establish executive-level reporting on delivery status, risks, and token ramp forecasts. Improve repeatability of partner onboarding, technical integration, and scaling motions. Manage escalations across internal and external stakeholders during high-severity delivery issues. Translate ambiguous infrastructure constraints into clear execution plans. Help define the long-term operating model for Token-as-a-Service across Stargate and 3P ecosystems. Qualifications 8+ years of Technical Program Management, Engineering Program Management, or Infrastructure Delivery experience. Experience leading large-scale technical programs involving cloud, data center, networking, hardware, or distributed systems. Strong understanding of compute infrastructure, clusters, networking, storage, and production systems. Proven ability to drive cross-functional execution across engineering, operations, finance, and external vendors. Experience managing executive stakeholders and communicating complex tradeoffs clearly. Strong analytical skills with ability to reason about utilization, throughput, capacity, and operational metrics. Comfortable operating in ambiguous, fast-scaling environments. Strong written and verbal communication skills. High ownership mentality with bias toward action. Experience working with external providers, strategic partners, or hyperscalers is highly preferred. Preferred Skills Experience with GPU clusters, AI infrastructure, or large-scale model serving environments. Familiarity with token economics, inference capacity planning, or workload scheduling. Experience scaling global infrastructure through third-party providers. Background in systems engineering, networking, or hardware deployment programs. Experience building new operational models in high-growth environments. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic. For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement . Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations. To notify OpenAI that you believe this job posting is non-compliant, please submit a report through this form . No response will be provided to inquiries unrelated to job posting compliance. We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link . OpenAI Global Applicant Privacy Policy At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.
About the Team OpenAI’s Infrastructure organization builds the systems that power frontier AI workloads at global scale. As compute demand accelerates, our ability to rapidly convert infrastructure investments into usable production capacity has become mission critical. The CPU / Storage / PoP / WAN team is responsible for the end-to-end infrastructure layers required to bring compute online: server and cluster activation, storage platforms, Points of Presence (PoPs), backbone connectivity, and global network expansion. We operate across first-party facilities, colocation environments, and strategic cloud partners to ensure OpenAI can scale reliably and quickly. About the Role We are seeking a highly technical Program Manager to lead execution across CPU, Storage, PoP, and WAN infrastructure programs that directly unlock OpenAI’s next generation compute capacity. In this role, you will own complex cross-functional programs spanning compute cluster activation, storage deployment, PoP bring-up, and backbone expansion. You will coordinate hardware readiness, site readiness, network pathing, storage availability, vendor execution, and engineering dependencies required to turn contracted infrastructure into live training and inference capacity. This role requires strong technical fluency across hardware systems, network infrastructure, storage architecture, and deployment execution. You should be comfortable operating from rack-level implementation details through executive-level capacity planning discussions. This role is based in San Francisco, CA, with travel as needed. Key Responsibilities Lead end-to-end execution of CPU / GPU cluster activation programs across OpenAI’s global infrastructure footprint Drive readiness to convert contracted compute capacity into schedulable production clusters Own deployment programs for new PoPs, backbone nodes, WAN expansion, and interconnection initiatives Build integrated schedules spanning procurement, logistics, installation, storage readiness, network turn-up, testing, and production handoff Coordinate BOM readiness, server delivery, racks, optics, cabling, storage hardware, and vendor milestones Partner with engineering teams to align compute, storage, and networking dependencies before cluster activation Manage deployment of storage systems supporting training and inference workloads, including readiness, validation, performance checks, and scaling plans Coordinate backbone capacity expansion, cross-connects, inter-region pathing, and cloud interconnect readiness with Azure and third-party providers Lead physical deployment execution including rack-and-stack, hardware bring-up, L1 validation, and site acceptance criteria Build repeatable deployment playbooks, dashboards, governance cadences, and operating mechanisms for scale Identify risks early across supply chain, site readiness, technical constraints, and vendor execution, then drive mitigation plans Communicate milestones, escalations, and capacity forecasts to senior leadership Qualifications 8+ years of experience in technical program management, infrastructure deployment, network deployment, or data center operations Strong experience delivering programs involving compute, storage, networking, or large-scale infrastructure systems Working knowledge of servers, clusters, storage arrays, routers, switches, optics, and structured cabling Experience owning cross-functional programs across engineering, operations, supply chain, and external vendors Strong understanding of deployment lifecycles from planning and procurement through production handoff Ability to reason across physical infrastructure execution and logical systems architecture dependencies Proven ability to build integrated schedules and drive accountability across multiple stakeholders Strong executive communication skills with experience managing critical escalations and leadership updates Comfortable operating in fast-moving environments with aggressive timelines and evolving priorities Highly analytical with strong problem-solving and execution instincts Preferred Skills Experience at a hyperscaler, cloud provider, AI infrastructure company, or global network operator Experience deploying GPU clusters, HPC systems, or large training environments Familiarity with distributed storage systems and high-performance data infrastructure Experience with PoP deployments, WAN backbone expansion, or global network buildouts Experience working across first-party, colo, and cloud environments Experience building repeatable infrastructure deployment systems in high-growth environments About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic. For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement . Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations. To notify OpenAI that you believe this job posting is non-compliant, please submit a report through this form . No response will be provided to inquiries unrelated to job posting compliance. We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link . OpenAI Global Applicant Privacy Policy At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.
About the team OpenAI’s Forward Deployed Engineering team partners with customers to turn research breakthroughs into production systems. We operate at the intersection of customer delivery and core platform development. About the role Forward Deployed Engineers (FDEs) lead complex end-to-end deployments of frontier models in production alongside our most strategic customers. You will own discovery, technical scoping, system design, build, and production rollout, partnering directly with customer engineering and domain teams. You will measure success through production adoption, measurable workflow impact, and eval-driven feedback that changes product and model roadmaps. You’ll work closely with our Product, Research, Partnerships, GRC, Security, and GTM teams. This role is based in Singapore. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. 50% travel is expected. In this role you will Own technical delivery across multiple deployments from first prototype to stable production. Build full-stack systems that deliver customer value and sharpen how we learn. Embed closely with customer teams, understand their needs, and guide adoption of what you build. Scope work, sequence delivery, and remove blockers early. Make trade-offs between scope, speed, and quality; adjust plans to protect delivery. Contribute directly in the code when progress or clarity depends on it. Codify working patterns into tools, playbooks, or building blocks that others can use. Share field feedback that helps Research and Product understand where the models succeed and where they can improve. Keep teams moving through clarity and follow-through. You might thrive in this role if you Bring 5+ years of engineering or technical deployment experience that includes customer-facing work. Have scoped and delivered complex systems in fast-moving or ambiguous environments. Write and review production-grade code across frontend and backend using Python, JavaScript, or comparable stacks. Have built or deployed systems powered by LLMs or generative models and understand how model behaviour affects product experience. Simplify complexity and make fast, sound decisions under pressure. Communicate clearly with engineers, product teams, and customer stakeholders. Spot risks early and adjust without slowing down. Model calm and judgment when the stakes are high. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic. For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement . Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations. To notify OpenAI that you believe this job posting is non-compliant, please submit a report through this form . No response will be provided to inquiries unrelated to job posting compliance. We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link . OpenAI Global Applicant Privacy Policy At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.
About the Role As a Director, Compute & Infrastructure FP&A, you will own and drive the monthly forecasting process for the Compute & Infrastructure org by partnering with various stakeholders across Finance, Accounting, Tax and Engineering. You will play a critical role in planning and forecasting the company’s largest and most complex cost center ( Compute & Infrastructure ). You will collaborate cross-functionally to develop long-range infrastructure investment plans, evaluate build vs. buy decisions, and ensure capital is deployed efficiently to support rapid growth. You will also provide strategic financial guidance through scenario modeling, ROI analysis, and performance tracking, enabling leadership to make high-stakes decisions under uncertainty. What You’ll Do Own compute financial planning & Forecasting. Build and manage consolidation models for GPU/CPU capacity, storage, networking, and data center investments. Translate infrastructure roadmaps into short- and long-term financial forecasts (LRP, annual planning) Coordinate closely with Corporate FP&A on timelines and process Present insights on a monthly basis to senior management. Drive infrastructure investment decisions. Evaluate build vs. buy, vendor vs. owned infrastructure, and capacity allocation tradeoffs. Develop frameworks for investment trade-offs to guide executive decision making. Build scalable tooling & reporting. Implement stakeholder-facing dashboards to track compute spend, utilization, and efficiency metrics. Improve visibility into unit economics (e.g., cost per training run, cost per inference, cost per customer). Drive forecasting accuracy & accountability. Lead budget vs. actual analysis for compute and infrastructure spend. Identify key cost drivers (utilization, pricing, efficiency gains) and reduce forecast variance. Support close & financial reporting. Partner with Accounting to ensure accurate classification of infrastructure spend (OpEx vs CapEx). Translate complex infrastructure costs into clear insights for leadership. Enable strategic decision-making. Build scenario models to support leadership decisions on capacity scaling, new model launches, and infrastructure investments. Lead ad hoc analyses on emerging topics. You Might Thrive in This Role If You Have 10+ years in strategic finance, with experience in infrastructure, cloud, hardware, or compute-intensive environments 2+ years in investment banking Must have experience running an FP&A team at the corporate level or business unit level with significant scale. Strong financial modeling skills, particularly in capacity planning, unit economics, and scenario analysis under uncertainty. Experience supporting large-scale infrastructure or cloud spend (e.g., AWS/GCP/Azure, GPUs, data centers). Ability to translate technical concepts (compute usage, model training/inference, system architecture) into financial insights. Proficiency in Excel/Sheets, SQL, and BI tools (e.g., Tableau); experience with planning systems like Anaplan is a plus. Strong cross-functional partnership skills, especially with Engineering, Product, and Supply Chain. Familiarity with AI/ML infrastructure cost drivers and the economics of training and serving models. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic. For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement . Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations. To notify OpenAI that you believe this job posting is non-compliant, please submit a report through this form . No response will be provided to inquiries unrelated to job posting compliance. We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link . OpenAI Global Applicant Privacy Policy At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.
About the Team ChatGPT is a rapidly evolving system: new capabilities ship continuously, product surfaces change quickly, and usage patterns shift week-to-week. Supporting that pace requires infrastructure that can handle real production constraints—high concurrency, unpredictable traffic patterns, complex dependency graphs, and frequent change. The ChatGPT Infrastructure team builds and operates the platforms that enable fast iteration without compromising performance or reliability. We design shared systems, data paths, rollout mechanisms, and reliability guardrails that teams rely on to ship changes to ChatGPT at scale. We focus on high-leverage infrastructure: primitives and “golden paths” that incorporate operational lessons as defaults, so engineers don’t need to rediscover failure modes, latency pitfalls, or integration issues each time they build something new. About the Role We’re hiring Senior and Staff Engineers to design and build infrastructure systems that underlie ChatGPT and multiply the effectiveness of teams building user experiences. This is not a support-only role. It’s a platform-building role: you’ll define interfaces, develop core abstractions, and create tooling to make safe, fast iteration the norm. Your work will reduce friction, prevent regressions, improve performance, and ensure systems scale gracefully as the product grows. Where You Can Have Impact You might work on one or more of the following areas (without being restricted to any single area): Platform foundations & frameworks: Core libraries, service frameworks, and shared components that standardize system building, integration, and evolution. Scalability & performance primitives: Patterns and infrastructure that reduce tail latency, improve throughput, and keep costs predictable as demand increases. Reliability guardrails: Mechanisms that prevent outages by design—rate limiting, load shedding, dependency isolation, backpressure, safe fallbacks, and robust regression controls. Developer productivity via golden paths: Paved roads for common workflows (data access patterns, service integration, request lifecycle) that are fast, safe, and easy to use. Observability & debugging systems: Instrumentation, metrics models, and investigative tooling that turn vague symptoms (“it’s slow”) into precise, actionable diagnoses. Safe change management: Deployment and rollout systems that support rapid iteration with confidence—progressive delivery, automated verification, and fast rollback strategies. Interface and contract design across boundaries: Clean APIs and stable contracts that reduce coupling and enable independent evolution across a complex ecosystem. What You’ll Do Build and evolve infrastructure platforms used by many engineers and services. Translate real-world constraints into clean abstractions: simple APIs, enforceable contracts, safe defaults. Drive improvements in reliability and performance through principled design, measurement, and iterative hardening. Partner with engineering and product teams to identify systemic pain points and develop reusable solutions. Own outcomes end-to-end: design → implementation → rollout → operational maturity. Qualifications Minimum Qualifications Experience building and operating large-scale distributed systems in production (high throughput, concurrency, and failure handling). Strong fundamentals in systems design, including caching, consistency, queueing/backpressure, and resilient dependency management. Ability to reason about performance (latency distributions, tail behavior, bottlenecks) and translate analysis into concrete engineering work. Track record of building platforms or shared infrastructure that improves velocity and correctness for other teams. Excellent communication and collaboration skills—aligning on interfaces, navigating tradeoffs, and driving cross-team execution. Preferred Qualifications Experience designing paved roads / golden paths (frameworks, libraries, self-serve tooling) that shape engineering behavior at scale. Deep understanding of reliability techniques: graceful degradation, circuit breakers, load shedding, rate limiting, and fault isolation. Experience building systems for safe iteration: progressive delivery, correctness checks, automated rollout gates, and production validation. Strong instincts for API and contract design—how to create interfaces that are stable, evolvable, and hard to misuse. Prior work that demonstrates “force multiplier” impact: enabling many teams via a small set of well-crafted primitives. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic. For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement . Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations. To notify OpenAI that you believe this job posting is non-compliant, please submit a report through this form . No response will be provided to inquiries unrelated to job posting compliance. We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link . OpenAI Global Applicant Privacy Policy At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.
About the Team OpenAI is building the infrastructure foundation for the next generation of AI. The Data Center Engineering team defines the strategy, reference architectures, technical requirements, and delivery standards for the large-scale data centers that support OpenAI research, products, and infrastructure partners. As a Data Center Controls Network Engineer, you will design, validate, and scale the controls and OT network architectures that support high-density AI data centers. You will work across controls systems, OT infrastructure, telemetry, commissioning, deployment, and operations, partnering with mechanical, electrical, IT/networking, security, and external delivery teams. About the Role We are seeking a mid to senior OT Network Engineer with a strong controls systems background to lead the design and operation of resilient, secure, and scalable OT network architectures for high-density AI data centers. This role translates compute, power, cooling, and operational requirements into practical OT network designs, evaluates vendor solutions, and drives technical decisions across controls infrastructure, telemetry, commissioning, and operations. The ideal candidate has strong hands-on experience in mission-critical OT environments, including industrial networking, virtualized infrastructure, and OT network operations, with expertise in routing, switching, segmentation, firewall policy, time synchronization, monitoring, and network lifecycle support. Key Responsibilities Define controls, automation, and OT network requirements for AI data center campuses. Develop reference architectures, engineering standards, and reusable design templates. Review and develop basis-of-design and functional design documents, including OT network diagrams, IP/VLAN schemes, telemetry architectures, data flow diagrams, and commissioning requirements. Design OT and infrastructure network architectures, including physical topology, logical topology, IP addressing, subnetting, VLANs, routing, switching, redundancy, segmentation, firewall policy coordination, out-of-band management, monitoring, and remote access patterns. Develop day-two network operations requirements, including change management, configuration backups, golden configurations, monitoring thresholds, firmware lifecycle, rollback plans, and post-change validation. Partner with electrical, mechanical, IT/networking, security, and operations teams to ensure OT network systems align with GPU deployments, campus-wide telemetry, and failure-domain isolation requirements. Define integration patterns and protocol requirements across BACnet/IP, BACnet MSTP, Modbus TCP/RTU, OPC UA, IEC-61850 MMS/GOOSE, MQTT, SNMP, syslog, NTP/PTP, IRIG-B, and vendor-specific interfaces. Lead technical evaluation of controls integrators, network equipment suppliers, design consultants, contractors, and commissioning agents Review network equipment submittals, configurations, firmware assumptions, certifications, test reports, and quality documentation. Support factory witnessed testing (FWT), site acceptance testing, network readiness checks, failover testing, and integrated systems testing. Troubleshoot complex controls network issues including packet loss, latency, duplicate IPs, routing errors, firewall drops, protocol incompatibilities, time synchronization drift, and intermittent device communication failures. Qualifications 8+ years of relevant experience in controls engineering, industrial automation, OT networking, mission-critical facilities, or similar critical infrastructure environments. Strong expertise in resilient OT network architecture, implementation, troubleshooting, and lifecycle support. Experience with OT/IT boundary design, secure enterprise integration, firewall policy design, redundant topologies, out-of-band management, and monitoring. Hands-on experience with Layer 3 OT network design, including IP addressing, subnetting, routing, VRFs, ACLs, inter-VLAN traffic control, and network segmentation. Hands-on experience with Layer 2 security and switching controls, including MACsec, port security, loop prevention, and switch-level access control. Hands-on experience in designing resilient OT network topologies using industrial redundancy protocols and architectures such as PRP, HSR, Cisco REP, RSTP/MSTP, and ring or star topologies. Hands-on experience in designing resilient infrastructure network architectures using HSRP/VRRP, spine-leaf topologies, redundant uplinks, and failure-domain isolation. Hands-on experience with industrial and infrastructure network equipment such as Cisco switches/routers, Juniper switches/routers, Palo Alto firewalls, Rockwell Automation Stratix switches, Siemens Ruggedcom or comparable industrial networking platforms. Experience with network management and observability platforms such as Cisco Catalyst Center (DNA Center), Palo Alto Panorama, Juniper Mist, industrial NMS tools, packet brokers, and OT monitoring platforms. Hands-on experience with industrial Ethernet, VPN tunneling, IPsec-based connectivity, and secure remote access. Hands-on experience with virtualized OT or controls server environments such as VMware vSAN, Microsoft Azure Stack HCI / Hyper-V, or comparable infrastructure platforms. Experience with industrial communication and OT infrastructure protocols, including BACnet/IP, BACnet MSTP, Modbus TCP/RTU, OPC UA, IEC-61850 MMS/GOOSE, MQTT, SNMP, syslog, NTP/PTP, IRIG-B, and vendor-specific interfaces, and strong understanding of their behavior across OT network architectures. Experience reviewing and producing technical design documentation, commissioning plans, and acceptance test procedures. Experience with factory witnessed testing, site acceptance testing, failover testing, telemetry validation, protocol compatibility testing, and root-cause analysis. Ability to use logs, packet captures, and field observations to make sound technical decisions and communicate risk clearly. Bachelor’s degree in Electrical Engineering, Computer Engineering, Network Engineering, Systems Engineering, or a related discipline. Preferred Skills Master's degree in Electrical Engineering, Computer Engineering, Network Engineering, Systems Engineering, or a related discipline. Experience leading multi-campus OT network integration, commissioning, and operations across cross-functional teams, contractors, vendors, and delivery partners. Relevant networking certifications such as Cisco CCNA/CCNP, Palo Alto PCNSA/PCNSE, Juniper JNCIA/JNCIS, or similar networking credentials. Cybersecurity certifications such as CISSP, GICSP, ISA/IEC 62443, CompTIA Security+, or similar cybersecurity credentials are a plus. Experience with network automation, Git-based configuration management, and Infrastructure as Code (IaC) using tools such as Ansible, Terraform, Python, or similar to support scalable OT network deployment and lifecycle management. Experience with scripting, APIs, and automation workflows that improve OT network operations. Experience using AI agents or MCP-connected tools to support telemetry analysis,and troubleshooting. Experience with relational database systems such as PostgreSQL, SQL Server, MySQL, or similar platforms used for OT telemetry, historian integrations, troubleshooting, and reporting. Work Environment and Travel This role requires periodic travel to data center campuses, vendors, labs, construction sites, commissioning activities, and controls/network cutovers. The engineer should be comfortable working across office, lab, construction, and live data center environments, including PPE, lockout/tagout, cyber hygiene, and change-control requirements. Work may include time-sensitive support during commissioning, startup, vendor testing, cutovers, network changes, telemetry issues, automation failures, and operational events. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic. For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement . Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations. To notify OpenAI that you believe this job posting is non-compliant, please submit a report through this form . No response will be provided to inquiries unrelated to job posting compliance. We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link . OpenAI Global Applicant Privacy Policy At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.
About the Team OpenAI’s Infrastructure organization builds and evaluates the systems that power advanced AI workloads. We work closely with hardware, modeling, and architecture teams to ensure that new platforms deliver real-world performance aligned with workload needs. Our team focuses on understanding workload behavior across evolving hardware platforms—bridging the gap between theoretical capability and observed system performance. About the Role We are seeking a Workload Porting & Performance Engineer to evaluate new hardware platforms by porting benchmarks and real-world workloads, analyzing performance, and identifying system bottlenecks. In this role, you will bring up workloads on new systems, characterize performance behavior, and adapt workloads to better utilize hardware capabilities. You will play a critical role in validating new platforms and ensuring that performance aligns with expectations across compute, memory, and networking subsystems. This role requires strong hands-on experience with performance analysis, workload optimization, and system-level debugging across hardware and software boundaries. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance. Key Responsibilities Port and enable benchmarks and real-world workloads on new hardware platforms. Evaluate system performance across compute, memory, storage, and networking subsystems. Identify and analyze performance bottlenecks and inefficiencies. Adapt and optimize workloads to better utilize hardware capabilities. Develop and run performance experiments and profiling workflows. Compare expected vs. observed performance and provide feedback to: hardware architecture teams performance modeling teams system and software engineers. Debug issues across the stack, including software, runtime, and hardware interactions. Provide actionable insights to guide platform readiness and deployment decisions. Qualifications Experience with performance analysis, benchmarking, or workload optimization. Strong understanding of system architecture, including CPU/GPU, memory, and I/O subsystems. Experience porting or adapting workloads across different hardware platforms. Familiarity with profiling tools and performance debugging techniques. Ability to identify root causes of performance issues across hardware/software boundaries. Experience working in large-scale or distributed system environments. Preferred Skills Experience with AI/ML workloads, including training or inference systems. Familiarity with GPU or accelerator-based systems. Experience working with low-level performance tools (profilers, tracing, microbenchmarks). Background in systems software, compilers, or runtime optimization. Experience collaborating with hardware and architecture teams on performance validation. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic. For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement . Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations. To notify OpenAI that you believe this job posting is non-compliant, please submit a report through this form . No response will be provided to inquiries unrelated to job posting compliance. We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link . OpenAI Global Applicant Privacy Policy At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.
About the Team OpenAI’s Hardware organization develops system and infrastructure solutions tailored to the demands of advanced AI workloads. We work across the full stack—from silicon to system integration—partnering closely with internal teams and external vendors to define and deliver next-generation AI infrastructure. Our team focuses on defining scalable, high-performance system architectures and reference designs that balance performance, cost, and operational efficiency across rapidly evolving technologies. About the Role We are seeking a 3P Architect to define and drive rack- and cluster-level reference designs in collaboration with external partners. This role is responsible for translating workload requirements and system-level goals into concrete architectures, aligning partners on critical design attributes, and ensuring vendor roadmaps meet our infrastructure needs. You will work closely with performance modeling and internal architecture teams to evaluate tradeoffs, while owning the end-to-end definition and execution of third-party system designs. This includes identifying gaps in current technologies, driving vendor development, and shaping future infrastructure capabilities. This role requires strong system intuition, cross-functional leadership, and the ability to operate effectively across internal teams and external ecosystems. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance. Key Responsibilities Define rack- and cluster-level reference architectures for AI infrastructure deployments. Translate workload requirements into clear system design specifications and partner deliverables. Collaborate with performance modeling teams to evaluate architectural tradeoffs and system behaviors. Align internal stakeholders and external partners on critical system attributes (performance, cost, power, reliability, scalability). Identify gaps in current technology offerings and drive vendors (ODM/JDM, silicon, networking) to close those gaps. Influence and shape vendor roadmaps to meet future infrastructure needs. Track emerging technologies and evaluate their applicability to AI systems. Define and lead proof-of-concept (PoC) efforts to validate new architectures and technologies. Act as a key interface between OpenAI and external partners, ensuring execution against design intent. Qualifications Have strong experience in system architecture for large-scale infrastructure or data center environments. Understand AI workload characteristics and how they map to system-level design decisions. Are comfortable working with performance modeling outputs to inform architectural direction. Have experience working with or managing hardware vendors (ODM/JDM, silicon, networking). Can drive alignment across multiple stakeholders with competing constraints. Have a track record of turning ambiguous requirements into clear, executable system designs. Are proactive in identifying gaps and driving solutions across organizational boundaries. Preferred Skills Experience defining rack- or cluster-level systems for hyperscale or AI workloads. Familiarity with accelerators (GPUs/ASICs), interconnects, and data center networking architectures. Experience influencing vendor roadmaps and reference designs. Background in infrastructure deployment, hardware engineering, or systems integration. Experience leading PoCs or early-stage hardware validation efforts. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic. For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement . Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations. To notify OpenAI that you believe this job posting is non-compliant, please submit a report through this form . No response will be provided to inquiries unrelated to job posting compliance. We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link . OpenAI Global Applicant Privacy Policy At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.
About the Team OpenAI is building the infrastructure foundation for the next generation of AI. The Data Center Engineering team defines the strategy, reference architectures, technical requirements, and delivery standards for the large-scale data centers that support OpenAI research, products, and infrastructure partners. As a Data Center Infrastructure Engineering Program Manager, you will help turn complex infrastructure strategy into executable programs across electrical, mechanical, controls, network, hardware, construction, commissioning, deployment, and operations workstreams. You will partner with research, hardware engineering, data center engineering, site development, supply chain, security, EHS, finance, legal, operations, and external delivery partners to bring OpenAI's infrastructure vision to life. About the Role We are looking for an Engineering Program Manager (EPM) to lead assigned infrastructure programs focused on production and non-production network integration, controls coordination, and the design and deployment of data hall or whitespace facilities. The EPM will support functional Directly Responsible Individuals (DRIs) across network, controls, structural, electrical, and mechanical disciplines. Key responsibilities include coordinating assigned workstreams and program controls, maintaining risks and interfaces, and supporting readiness within the network and data hall deployment track. The ideal candidate thrives on bringing structure to complex environments characterized by ambiguous technical requirements, large partner ecosystems, tight deadlines, and high operational stakes. This individual must be adept at keeping teams aligned on decisions, risks, dependencies, schedules, and readiness criteria, and escalating gaps or decision points when needed. Candidates should have a proven track record of managing technically challenging engineering programs across major lifecycle phases, including design, validation, procurement, construction, commissioning, deployment, and operational handoff. Key Responsibilities Translate assigned infrastructure goals into clear workstream charters, scopes, milestones, owners, decision points, success metrics, resourcing assumptions, and execution plans. Build and maintain integrated execution plans for assigned programs covering network, controls, data hall design, whitespace deployment, commissioning preparation, and deployment readiness. Support coordination across network, controls, structural, electrical, mechanical, hardware integration, construction, commissioning, and operations teams. Work with third-party design teams to define design milestones from concept through detailed design, including basis-of-design development, requirements tracking, design reviews, technical comment resolution, change management, and release readiness. Maintain the dependency map, issue log, risk register, action tracker, and decision log for assigned network and data hall workstreams. Coordinate design and review milestones for network rooms, non-production network services, OT / IT interface points, rack deployment assumptions, telemetry interfaces, controls dependencies, and data hall deployment packages. Track building-level network and support-space interfaces such as MPOE, MMR, Network Core, WAN, support rooms, and associated handoff points where they affect assigned programs. Support the network and controls DRIs by organizing reviews, resolving cross-discipline gaps, surfacing decisions, and keeping partner deliverables aligned to schedule. Manage partner and vendor deliverables such as submittals, interface packages, installation assumptions, turn-up plans, readiness evidence, field issue logs, and corrective action tracking. Drive readiness tracking for assigned 1P, 3P, colo, and selected CSP programs, including bring-up sequencing, installation readiness, access dependencies, maintenance windows, and first-use criteria. Prepare clear status updates, dashboards, and executive-ready summaries for the Industrial Compute lead and project stakeholders. Capture lessons learned from deployment and handoff activities and feed them back into playbooks, standards, and interface definitions. Qualifications Extensive experience in engineering program management, technical program management, mission-critical infrastructure delivery, data center deployment, or comparable complex execution environments, typically gained through 10+ years of relevant work or equivalent depth of experience. Proven ability to operate within ambiguous, cross-functional engineering programs with shifting requirements, urgent timelines, and high-stakes operational risk, driving from concept through design, validation, procurement, construction, commissioning, deployment, and operations. Proven experience coordinating cross-functional programs that include network, controls, mechanical, electrical, structural, construction, commissioning, or operations participants. Strong technical fluency in at least several of the following areas: data hall deployment, non-production network, production network interfaces, controls coordination, telemetry, rack deployment, mission-critical support spaces, and infrastructure handoff. Experience building and maintaining integrated schedules, dependency maps, risk registers, decision logs, readiness trackers, and partner action plans. Experience coordinating external partners, vendors, design firms, delivery teams, or operators in a multi-party infrastructure environment. Ability to understand complex technical tradeoffs, ask strong questions, identify hidden dependencies, and help teams move toward clear decisions without needing to be the sole technical owner. Excellent written and verbal communication skills, with the ability to produce crisp status reporting and drive action across matrixed teams. Bachelor's degree in Engineering, Computer Science, Construction Management, Operations, Business, or a related technical or quantitative field, or equivalent practical experience. Preferred Skills Direct experience with hyperscale data centers, AI infrastructure, HPC environments, colocation, or partner-delivered data center programs. Experience supporting high-density data hall or HAC-related design and deployment efforts. Experience with non-production network and support-service readiness for large-scale infrastructure deployments. Experience coordinating controls integration, network-room layouts, telemetry interfaces, rack deployment packages, or early operational handoff. Comfort with technical documentation such as one-line diagrams, P&IDs, controls sequences, network diagrams, equipment specifications, interface control documents, telemetry schemas, test procedures, and commissioning scripts. Familiarity with 1P, 3P, colocation, and cloud service provider delivery models and the differing owner-partner interface expectations in each. Work Environment and Travel This role may require periodic travel to data center campuses, manufacturing partners, equipment suppliers, laboratories, construction sites, commissioning activities, and partner program reviews. The program manager should be comfortable working across office, lab, manufacturing, construction, and operating data center environments, including environments that require PPE, safety briefings, change-control discipline, and coordination with site operations. Work may include time-sensitive escalations during design reviews, procurement, manufacturing validation, commissioning, startup, production deployment, vendor testing, operational readiness, or operational incidents. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic. For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement . Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations. To notify OpenAI that you believe this job posting is non-compliant, please submit a report through this form . No response will be provided to inquiries unrelated to job posting compliance. We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link . OpenAI Global Applicant Privacy Policy At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.
About the Team OpenAI’s Hardware organization develops system and infrastructure solutions designed for the unique demands of advanced AI workloads. We work closely with architecture, infrastructure, and vendor teams to evaluate system performance and guide critical design decisions. Our team focuses on building and applying performance modeling frameworks to understand system behavior, quantify tradeoffs, and support next-generation infrastructure design. About the Role We are seeking an Performance Modeling Engineer to support the development and application of modeling tools used to evaluate AI system performance and inform architectural decisions. In this role, you will partner closely with Senior Performance Modeling Engineers and the Performance Modeling Lead to analyze system behavior, run simulations and analytical models, and help evaluate tradeoffs across compute, memory, networking, and storage. You will contribute to building modeling frameworks while developing a strong foundation in system architecture and AI infrastructure. This role is ideal for early-career engineers with 1–2 years of experience in software engineering, systems analysis, or performance modeling who are excited to grow in large-scale infrastructure and hardware/software systems. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance. Key Responsibilities Support the development and maintenance of performance modeling tools and frameworks Assist in building models to evaluate system behavior across compute, memory, networking, and interconnect subsystems Help analyze distributed system scaling behavior and identify performance bottlenecks Run simulations and analytical models to support architecture and infrastructure decisions Partner with senior engineers to evaluate design tradeoffs across hardware and system components Interpret modeling outputs and help translate findings into clear recommendations Validate models using benchmarking data and real system performance measurements Improve modeling workflows, documentation, and usability for broader team adoption Collaborate cross-functionally with hardware, infrastructure, and architecture teams Continuously build technical depth across AI infrastructure, system architecture, and performance analysis Qualifications 1–2 years of experience in software engineering, systems modeling, performance analysis, or related technical work Strong programming skills and experience building technical tools, scripts, or frameworks Familiarity with system architecture fundamentals such as compute, memory, and networking Ability to reason about system performance, bottlenecks, and scaling behavior Strong analytical and problem-solving skills with comfort working in quantitative environments Ability to learn quickly and work effectively across technical teams Preferred Skills Exposure to AI/ML workloads, distributed systems, or large-scale infrastructure Experience with simulation tools, benchmarking, profiling, or performance analysis Familiarity with data center systems, server architecture, or hardware platforms Interest in system architecture and hardware/software co-design Internship or early professional experience in performance engineering, infrastructure, or systems design About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic. For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement . Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations. To notify OpenAI that you believe this job posting is non-compliant, please submit a report through this form . No response will be provided to inquiries unrelated to job posting compliance. We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link . OpenAI Global Applicant Privacy Policy At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.
About the Team OpenAI’s Hardware organization develops system and infrastructure solutions designed for the unique demands of advanced AI workloads. We work closely with architecture, infrastructure, and vendor teams to evaluate system performance and guide critical design decisions. Our team focuses on building and applying performance modeling frameworks to understand system behavior, quantify tradeoffs, and inform next-generation infrastructure design. About the Role We are seeking Performance Modeling Engineers to develop and apply modeling tools that evaluate AI system performance and inform architectural decisions. In this role, you will work closely with the Performance Modeling Lead and partner teams to analyze system behavior, run simulations or analytical models, and help quantify tradeoffs across compute, memory, networking, and storage. You will contribute to building modeling frameworks and applying them to real-world questions that impact system design and vendor decisions. This role is well-suited for engineers with strong software or modeling backgrounds who are interested in developing deeper expertise in system architecture and AI infrastructure. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance. Key Responsibilities Develop and maintain performance modeling tools and frameworks. Build models to evaluate system behavior across: compute, memory, and interconnect subsystems distributed system scaling and bottlenecks. Run simulations and analytical models to support architectural tradeoff analysis. Collaborate with performance modeling lead and system architects to answer forward-looking design questions. Analyze and interpret modeling outputs, translating results into actionable insights. Validate models against real system measurements and workload behavior. Contribute to improving modeling fidelity, usability, and scalability. Qualifications Strong software engineering or modeling background (e.g., simulation, systems modeling, or performance analysis). Familiarity with system architecture fundamentals (compute, memory, networking). Experience with programming and building technical tools or frameworks. Ability to reason about performance bottlenecks and scaling behavior. Strong analytical skills and comfort working with quantitative models. Ability to collaborate across teams and learn new system domains quickly. Preferred Skills Exposure to AI/ML workloads or distributed systems. Experience with simulation tools, performance modeling, or systems analysis. Familiarity with data center infrastructure or large-scale systems. Experience working with performance data, benchmarking, or profiling tools. Interest in system architecture and hardware/software co-design. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic. For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement . Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations. To notify OpenAI that you believe this job posting is non-compliant, please submit a report through this form . No response will be provided to inquiries unrelated to job posting compliance. We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link . OpenAI Global Applicant Privacy Policy At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.