
Explore active software engineering, data, AI, product, and design roles directly from verified employers—and make sure your resume is ready before applying.
Get an instant ATS score, missing keyword alert, and bullet rewrites tailored to your target job before submitting your application.
About the team OpenAI’s Forward Deployed Engineering team partners with customers to turn research breakthroughs into production systems. We operate at the intersection of customer delivery and core platform development. About the Role This is a founding role. As a Technical Deployment Lead (TDL), you will define how OpenAI delivers complex systems to customers. You will own how they are built, shipped, and adopted. You’ll translate business outcomes into a technical plan, run day-to-day execution across FDEs, Researchers, and Customer Engineers, and partner with customer teams to ensure delivery supports their goals. This is not a management role, however you'll own delivery end-to-end: embedding with customers to map workflows and success criteria, ensuring components ship on time, and leading readiness and change management for adoption. You’ll track progress, manage dependencies, make sequencing decisions, and drive 0→1 prototypes through MVP and scale. You will also share field insights with Product and Research to guide roadmap and priorities. Success will be measured first and foremost by impact - deployments that deliver measurable value against customer goals, drive adoption, and become critical to their workflows. Additional measures of success include delivery reliability (milestones hit, low reopen/churn), operating leverage (patterns reused across deployments), judgment under pressure, and product impact (field signal that shifts roadmaps/architectures). This is a high-trust, high-autonomy role. Success requires deep technical project management expertise, extreme ownership of outcomes, and an ability to immerse in customer workflows and partner with customer teams to solve complex engineering problems at pace. This role is based in Tokyo. We use a hybrid model of 3 days in office and offer relocation assistance. Travel up to 25-50% is required. To succeed in this position you must be fully bilingual—fluent in both Japanese and English (spoken and written). Please note that your resume must be submitted in English, and the interview process will include conversations in both languages. In this role, you will: Own the technical delivery plan for multiple interdependent workstreams. Translate business objectives into a roadmap with milestones, dependencies, and acceptance criteria. Run day-to-day engineering execution. Track and drive delivery across OpenAI FDE and customer teams. Keep progress unblocked and sequenced. Make real-time trade-offs on scope and priority to protect the critical path. Embed with customer teams to land production deployments and drive adoption. Map workflows, shape tools/integrations, and translate requirements into a delivery plan. Lead onboarding, adoption, and change management. Partner with Product and Research so platform components and research workstreams land in time to support deployment goals. Codify solution patterns and evals. Extract reusable patterns and package field signals to improve product and models. Own value cases and ROI. Set impact hypotheses, baselines, and KPIs; run pre-/post-deployment measurement and report to exec sponsors. You’ll thrive in this role if you: Bring 7+ years of customer‑facing technical delivery leadership. Track record of successfully leading large, complex, high-stakes customer engagements where customer outcomes depended on tight coordination and fast decision making, ideally involving AI. Excel in high ambiguity environments. Know how to simplify complex and dynamic work. Move fluidly between system level understanding and execution level detail; can dive into customer workflows/data, map constraints, sketch architectures and move ambiguous problems to shipped systems. Think strategically and pattern-match. Able to step back from execution detail, recognise broader trends across deployments, and connect customer needs to scalable, reusable solutions. Have strong technical fluency and sharp sequencing instincts. Confident discussing technical details, pressure-testing architectures, and making trade-offs. Have shipped AI/LLM systems. You understand solution patterns, integration basics, and production pitfalls. You’re a translator with executive presence. You make complex technical trade-offs legible to business leaders and convert strategy into day-to-day technical execution. Enjoy being onsite with customers to accelerate delivery (often 25-50%, sometimes higher). Have expertise in at least one major sector (e.g., healthcare, energy, financial services, semiconductors, IT) to elevate solution framing and credibility. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic. For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement . Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations. To notify OpenAI that you believe this job posting is non-compliant, please submit a report through this form . No response will be provided to inquiries unrelated to job posting compliance. We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link . OpenAI Global Applicant Privacy Policy At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.
Who We Are Notion is the collaborative AI workspace where teams and agents think together . We're building one place where your knowledge, projects, meetings, and AI tools live side by side, so work is faster, clearer, and less fragmented. Millions of individuals, small teams, and large companies run their work on Notion. Notinos (our employees) are customer zero in bringing this future of work to life. We care about craft, building things that last, and the belief that great work is still fundamentally human. Our goal isn’t to ship the next feature. Each and every team of Notinos is working to set the standard for how humans work together in the AI era. From building a business’s system of record to making and managing AI agents to automating away the busy work, we care deeply about giving our customers more time for their life’s work. About The Role Notion is rebuilding its go-to-market motion around a workflow-first, builder-led model—and the BDR is at the front of it. As a Business Development Rep, you'll be the force that turns the right accounts into the right conversations. You'll partner closely with Solutions Consultants to identify and target companies that are ready to rethink how they work—and your job is to get them in the room. This isn't generic pipeline work: you're qualifying for workflow-first, build-ready opportunities that set SCs up to win. If you're competitive, curious, and energized by outbound that actually means something, this is the role for you. What You'll Achieve Generate qualified, workflow-first pipeline by identifying and targeting high-potential accounts in your assigned book, partnering with your sellers on account and outbound strategy Run multi-channel outbound campaigns—email, phone, LinkedIn, and events—to create urgency and secure first meetings with the right decision-makers Conduct initial qualification to ensure meetings handed to sellers are "ready to build," not just "ready to talk"—validating ICP, intent, and workflow fit before the handoff Document and deliver clean handoff packets in Salesforce with ICP signals, pain points, workflow hypothesis, and stakeholder context so sellers can hit the ground running Leverage GTM System tooling and intent data to sharpen targeting, prioritize accounts by signal strength, and increase the quality and consistency of your outbound motion Skills You'll Need to Bring Fluency in Korean and business level in English Proven track record in outbound pipeline generation, with hands-on experience in cold prospecting across phone, email, and social channels Strong written and verbal communication skills with the ability to craft compelling, persona-specific outreach that earns a response Sharp qualification instincts—able to quickly assess ICP fit, identify urgency, and set solutions consultants up for high-quality first conversations Passion for AI/ML and understanding of API-first or consumption-based business models High-activity mindset with the discipline to prioritize based on signal and data, not just volume You don't need to be an AI expert, but you're curious and willing to adopt AI tools to work smarter and deliver better results. Nice to Haves Experience with Salesforce, Outreach, or similar GTM tooling Familiarity with Notion as a product or prior experience selling productivity, collaboration, or workflow software Background in a high-velocity, metric-driven outbound environment with clear pipeline targets By clicking “Submit Application”, I understand and agree that Notion and its affiliates and subsidiaries will collect and process my information in accordance with Notion’s Global Recruiting Privacy Policy . #LI-Onsite A Note on AI You don’t need deep AI expertise for every role, but we do expect every Notino to be intellectually curious, drawn to tinkering and discovery, and excited to use AI as a real collaborator in their work. For some roles, AI fluency is a core requirement — when that’s the case, we'll say so explicitly in the qualifications. People who thrive here don’t treat AI as a novelty. They use it to think better, and make their work easier for others to build on. Equal Opportunity & Accommodations We hire talented people from a wide range of backgrounds. If you’re excited about this role but don’t meet every bullet, we still encourage you to apply. Notion is an equal opportunity employer and does not discriminate on the basis of any legally protected characteristic. Consistent with applicable law, we will consider for employment qualified applicants with arrest and conviction records. Notion provides reasonable accommodations during the application process; if you need one, please let your recruiter know. Notion is proud to be an equal opportunity employer. We do not discriminate in hiring or any employment decision based on race, color, religion, national origin, age, sex (including pregnancy, childbirth, or related medical conditions), marital status, ancestry, physical or mental disability, genetic information, veteran status, gender identity or expression, sexual orientation, or other applicable legally protected characteristic. Notion considers qualified applicants with criminal histories, consistent with applicable federal, state and local law. Notion is also committed to providing reasonable accommodations for qualified individuals with disabilities and disabled veterans in our job application procedures. If you need assistance or an accommodation due to a disability, please let your recruiter know.
Who We Are Notion is the collaborative AI workspace where teams and agents think together . We're building one place where your knowledge, projects, meetings, and AI tools live side by side, so work is faster, clearer, and less fragmented. Millions of individuals, small teams, and large companies run their work on Notion. Notinos (our employees) are customer zero in bringing this future of work to life. We care about craft, building things that last, and the belief that great work is still fundamentally human. Our goal isn’t to ship the next feature. Each and every team of Notinos is working to set the standard for how humans work together in the AI era. From building a business’s system of record to making and managing AI agents to automating away the busy work, we care deeply about giving our customers more time for their life’s work. About The Role Notion is rebuilding its go-to-market motion around a workflow-first, builder-led model—and the BDR is at the front of it. As a Business Development Rep, you'll be the force that turns the right accounts into the right conversations. You'll partner closely with Solutions Consultants to identify and target companies that are ready to rethink how they work—and your job is to get them in the room. This isn't generic pipeline work: you're qualifying for workflow-first, build-ready opportunities that set SCs up to win. If you're competitive, curious, and energized by outbound that actually means something, this is the role for you. What You'll Achieve Generate qualified, workflow-first pipeline by identifying and targeting high-potential accounts in your assigned book, partnering with your sellers on account and outbound strategy Run multi-channel outbound campaigns—email, phone, LinkedIn, and events—to create urgency and secure first meetings with the right decision-makers Conduct initial qualification to ensure meetings handed to sellers are "ready to build," not just "ready to talk"—validating ICP, intent, and workflow fit before the handoff Document and deliver clean handoff packets in Salesforce with ICP signals, pain points, workflow hypothesis, and stakeholder context so sellers can hit the ground running Leverage GTM System tooling and intent data to sharpen targeting, prioritize accounts by signal strength, and increase the quality and consistency of your outbound motion Skills You'll Need to Bring Proven track record in outbound pipeline generation, with hands-on experience in cold prospecting across phone, email, and social channels Strong written and verbal communication skills with the ability to craft compelling, persona-specific outreach that earns a response Sharp qualification instincts—able to quickly assess ICP fit, identify urgency, and set solutions condsultants up for high-quality first conversations Passion for AI/ML and understanding of API-first or consumption-based business models High-activity mindset with the discipline to prioritize based on signal and data, not just volume You don't need to be an AI expert, but you're curious and willing to adopt AI tools to work smarter and deliver better results. Nice to Haves Experience with Salesforce, Outreach, or similar GTM tooling Familiarity with Notion as a product or prior experience selling productivity, collaboration, or workflow software Background in a high-velocity, metric-driven outbound environment with clear pipeline targets Notion is committed to providing highly competitive cash compensation, equity, and benefits. The compensation offered for this role will be based on multiple factors such as location, the role’s scope and complexity, and the candidate’s experience and expertise, and may vary from the range provided below. For this role, based in New York City, the estimated hourly rate is $33.65 - $38.70 per hour with a 30k annual commission target, annualized to salary range of $100,000 - $115,000 per year. By clicking “Submit Application”, I understand and agree that Notion and its affiliates and subsidiaries will collect and process my information in accordance with Notion’s Global Recruiting Privacy Policy and NYLL 144 . #LI-Onsite A Note on AI You don’t need deep AI expertise for every role, but we do expect every Notino to be intellectually curious, drawn to tinkering and discovery, and excited to use AI as a real collaborator in their work. For some roles, AI fluency is a core requirement — when that’s the case, we'll say so explicitly in the qualifications. People who thrive here don’t treat AI as a novelty. They use it to think better, and make their work easier for others to build on. Equal Opportunity & Accommodations We hire talented people from a wide range of backgrounds. If you’re excited about this role but don’t meet every bullet, we still encourage you to apply. Notion is an equal opportunity employer and does not discriminate on the basis of any legally protected characteristic. Consistent with applicable law, we will consider for employment qualified applicants with arrest and conviction records. Notion provides reasonable accommodations during the application process; if you need one, please let your recruiter know. Notion is proud to be an equal opportunity employer. We do not discriminate in hiring or any employment decision based on race, color, religion, national origin, age, sex (including pregnancy, childbirth, or related medical conditions), marital status, ancestry, physical or mental disability, genetic information, veteran status, gender identity or expression, sexual orientation, or other applicable legally protected characteristic. Notion considers qualified applicants with criminal histories, consistent with applicable federal, state and local law. Notion is also committed to providing reasonable accommodations for qualified individuals with disabilities and disabled veterans in our job application procedures. If you need assistance or an accommodation due to a disability, please let your recruiter know.
Who We Are Notion is the collaborative AI workspace where teams and agents think together . We're building one place where your knowledge, projects, meetings, and AI tools live side by side, so work is faster, clearer, and less fragmented. Millions of individuals, small teams, and large companies run their work on Notion. Notinos (our employees) are customer zero in bringing this future of work to life. We care about craft, building things that last, and the belief that great work is still fundamentally human. Our goal isn’t to ship the next feature. Each and every team of Notinos is working to set the standard for how humans work together in the AI era. From building a business’s system of record to making and managing AI agents to automating away the busy work, we care deeply about giving our customers more time for their life’s work. About The Role Notion is rebuilding its go-to-market motion around a workflow-first, builder-led model—and the BDR is at the front of it. As a Business Development Rep, you'll be the force that turns the right accounts into the right conversations. You'll partner closely with Solutions Consultants to identify and target companies that are ready to rethink how they work—and your job is to get them in the room. This isn't generic pipeline work: you're qualifying for workflow-first, build-ready opportunities that set SCs up to win. If you're competitive, curious, and energized by outbound that actually means something, this is the role for you. What You'll Achieve Generate qualified, workflow-first pipeline by identifying and targeting high-potential accounts in your assigned book, partnering with your sellers on account and outbound strategy Run multi-channel outbound campaigns—email, phone, LinkedIn, and events—to create urgency and secure first meetings with the right decision-makers Conduct initial qualification to ensure meetings handed to sellers are "ready to build," not just "ready to talk"—validating ICP, intent, and workflow fit before the handoff Document and deliver clean handoff packets in Salesforce with ICP signals, pain points, workflow hypothesis, and stakeholder context so sellers can hit the ground running Leverage GTM System tooling and intent data to sharpen targeting, prioritize accounts by signal strength, and increase the quality and consistency of your outbound motion Skills You'll Need to Bring Proven track record in outbound pipeline generation, with hands-on experience in cold prospecting across phone, email, and social channels Strong written and verbal communication skills with the ability to craft compelling, persona-specific outreach that earns a response Sharp qualification instincts—able to quickly assess ICP fit, identify urgency, and set solutions condsultants up for high-quality first conversations Passion for AI/ML and understanding of API-first or consumption-based business models High-activity mindset with the discipline to prioritize based on signal and data, not just volume You don't need to be an AI expert, but you're curious and willing to adopt AI tools to work smarter and deliver better results. Nice to Haves Experience with Salesforce, Outreach, or similar GTM tooling Familiarity with Notion as a product or prior experience selling productivity, collaboration, or workflow software Background in a high-velocity, metric-driven outbound environment with clear pipeline targets Notion is committed to providing highly competitive cash compensation, equity, and benefits. The compensation offered for this role will be based on multiple factors such as location, the role’s scope and complexity, and the candidate’s experience and expertise, and may vary from the range provided below. For this role, based in San Francisco the estimated hourly rate is $33.65 - $38.70 per hour with a 30k annual commission target, annualized to salary range of $100,000 - $115,000 per year. By clicking “Submit Application”, I understand and agree that Notion and its affiliates and subsidiaries will collect and process my information in accordance with Notion’s Global Recruiting Privacy Policy . #LI-Onsite A Note on AI You don’t need deep AI expertise for every role, but we do expect every Notino to be intellectually curious, drawn to tinkering and discovery, and excited to use AI as a real collaborator in their work. For some roles, AI fluency is a core requirement — when that’s the case, we'll say so explicitly in the qualifications. People who thrive here don’t treat AI as a novelty. They use it to think better, and make their work easier for others to build on. Equal Opportunity & Accommodations We hire talented people from a wide range of backgrounds. If you’re excited about this role but don’t meet every bullet, we still encourage you to apply. Notion is an equal opportunity employer and does not discriminate on the basis of any legally protected characteristic. Consistent with applicable law, we will consider for employment qualified applicants with arrest and conviction records. Notion provides reasonable accommodations during the application process; if you need one, please let your recruiter know. Notion is proud to be an equal opportunity employer. We do not discriminate in hiring or any employment decision based on race, color, religion, national origin, age, sex (including pregnancy, childbirth, or related medical conditions), marital status, ancestry, physical or mental disability, genetic information, veteran status, gender identity or expression, sexual orientation, or other applicable legally protected characteristic. Notion considers qualified applicants with criminal histories, consistent with applicable federal, state and local law. Notion is also committed to providing reasonable accommodations for qualified individuals with disabilities and disabled veterans in our job application procedures. If you need assistance or an accommodation due to a disability, please let your recruiter know.
About the Team Security is foundational to OpenAI’s mission to ensure that artificial general intelligence benefits all of humanity. The Security organization protects OpenAI’s technology, people, and products by building and operating deeply technical systems that must work reliably at massive scale. Our work underpins OpenAI’s commitments around safety, privacy, and security across research, products, and emerging platforms. The Host Assurance team exists to make bare metal a dependable, scalable foundation for OpenAI: secure by default, verifiable in practice, and resilient across providers and operating models. We operate at the trust boundary between physical hardware and cloud-scale orchestration, ensuring that hosts are eligible to safely run workloads with predictable security properties and auditability. About the Role OpenAI is seeking a Security Engineer, Host Assurance to help build the trust foundations for bare-metal platforms across OpenAI’s global infrastructure. This is a deeply hands-on engineering role for a builder who can design, implement, and operate the core security infrastructure that establishes trust in hardware platforms before they are eligible to run workloads. Success in this role requires strong technical judgment, the ability to work comfortably at low levels of the stack, and a practical mindset for building systems that are secure, reliable, and usable in fast-moving production environments. The systems you build will sit on the critical path of OpenAI’s frontier infrastructure investments and will directly shape how large amounts of compute are brought online - securely, responsibly, and at global scale - underpinning long-lived commitments around privacy, security, and reliability. You will partner closely with infrastructure, research, and confidential computing initiatives—including novel hardware platforms and emerging deployment models– to make the secure path the easiest path. This role is well suited for engineers who enjoy working across trust services, operating systems, hardware and firmware validation, and infrastructure security, and who are excited by ambiguous, high-impact problems at the boundary of hardware and large-scale systems. In this role, you will: Design, build, and operate components of the Host Assurance platform that establish trust in bare-metal hosts before they are eligible for production use. Help ensure hosts are verifiably trustworthy from delivery and installation through secure bootstrap and readiness to join orchestration systems. Build and improve systems such as machine identity, certificate issuance and enrollment, HSM-backed or key-management-backed trust services, host attestation, measurement, and baseline verification tooling. Validate delivered hardware and firmware against vendor claims and continuously detect and manage drift over time. Eliminate insecure bootstrap patterns while preserving deployment throughput and operational reliability. Partner with provisioning, fleet, and orchestration teams to deliver paved paths where the secure approach is the easiest approach. Contribute code, reviews, operational improvements, and design guidance for foundational trust services that must be dependable at scale. Help define observable, testable security properties for host platforms and improve the telemetry and validation needed to enforce them in practice. Participate in incident response, debugging, and post-incident improvements for security-critical infrastructure. Work across different deployment models and provider boundaries while maintaining a consistent bar for host trust outcomes. You might thrive in this role if you: Have strong software engineering experience building and operating reliable production systems at scale. Have deep expertise in at least one relevant domain such as PKI, HSMs, machine identity, applied cryptography, secure boot, firmware or hardware security, host attestation, or low-level platform security. Are comfortable working across systems boundaries, from services and APIs down to host, boot, firmware, or hardware-adjacent trust mechanisms. Can write production quality code and reason clearly about failure modes, operational safety, and long-term maintainability. Have experience replacing fragile or manual security mechanisms with durable, paved-path infrastructure. Balance rigor with pragmatism, and care about making strong security controls deployable in real-world environments. Are self-directed, low ego, and willing to work across disciplines to solve the most important problems. Enjoy building in ambiguous spaces where the architecture is still emerging, stakes are at all time high, and the future is being built. Workplace & Location This role is preferably based in San Francisco, CA or Seattle, WA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic. For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement . Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations. To notify OpenAI that you believe this job posting is non-compliant, please submit a report through this form . No response will be provided to inquiries unrelated to job posting compliance. We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link . OpenAI Global Applicant Privacy Policy At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.
About the Team The Scaling team is responsible for the architectural and engineering backbone of OpenAI’s infrastructure. We design and deliver advanced systems that support the deployment and operation of cutting-edge AI models. Our work spans system software, networking, platform architecture, fleet-level monitoring, and performance optimization. About the Role We’re hiring an SW Engineer to enable production workloads and end-to-end testing on new platforms. This role will include creating new test harnesses and platform stress benchmarks, porting existing inference and training workloads to new, sometimes early-access, systems/hardware, analyzing performance and bottlenecks, and characterizing the end-to-end behavior of new systems (compute, comms, storage, control plane, and failure modes). Key Responsibilities Port and validate key inference and training workloads on new platforms/SKUs as they arrive; drive correctness, performance, and stability to an internal readiness bar. Build a suite of benchmarks and stress tests that capture real E2E behavior of our workloads by exercising all aspects of a system, including CPU, GPU, memory subsystem, frontend, scale-up, and scale-out networking (including WAN traffic, NVlink and RDMA collectives), storage, thermals, and any other relevant parts. Deep-dive performance on distributed training/inference: Collective performance and tuning (across NCCL/RCCL and internal libraries) Overlap of compute/communication, kernel-level bottlenecks, memory bandwidth and scheduling effects Create repeatable test harnesses that run in CI / lab environments and produce actionable outputs (pass/fail, performance score, regression detection). Partner with systems + fleet bring-up engineers to ensure the platform is not only stable and performant, but also operationally usable and scalable (containerization, K8s integration, telemetry hooks, failure triage loops). Work cross-functionally with vendors and internal stakeholders by producing clear bug reports, minimal repros, and prioritized issue lists. Qualifications BS in CS/EE (or equivalent practical experience). 5+ years in one or more of: ML systems, performance engineering, distributed systems, or HPC. Strong hands-on experience with: PyTorch and modern LLM training/inference stacks Large-scale distributed training concepts (data/model/pipeline parallel, collective comms) Experience with RDMA and debugging/optimizing comms libraries (NCCL or RCCL) and their interaction with hardware/network Proficiency in Python plus comfort reading/writing performance-critical code (C++/CUDA/HIP is a plus). Strong profiling/debugging skills (e.g., Nsight, rocprof, perf, flamegraphs; ability to reason from traces/counters). Preferred Skills Experience building workload-shaped benchmarks and stress/fault tests that correlate to production behavior (not just synthetic loops or microbenchmarks). Familiarity with RDMA networking and transport tuning; understanding of how network topology and congestion impact collectives. Experience running and validating workloads in Kubernetes, and bridging “research code” into robust, repeatable infrastructure. Hands-on lab experience with early hardware (new NICs, new GPUs/accelerators, early racks). About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic. For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement . Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations. To notify OpenAI that you believe this job posting is non-compliant, please submit a report through this form . No response will be provided to inquiries unrelated to job posting compliance. We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link . OpenAI Global Applicant Privacy Policy At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.
About the Team OpenAI, in close collaboration with our capital partners, is embarking on a journey to build the world’s most advanced AI infrastructure ecosystem. The Infrastructure team is central to this mission, setting the core strategy and implementing the vision. From site selection to deployment to operations, this team sits at the intersection of commercial, technical, and operational domains, interacting with experts and executives inside and outside of OpenAI. We design and operate mission-critical facilities that support cutting-edge AI workloads at scale. About the Role We are seeking a Critical Facilities Lead to support the commissioning, deployment, and long-term operation of our next-generation AI data centers. This role bridges the interface between data center construction and hardware landing, ensuring seamless integration of mission-critical infrastructure with hardware deployment timelines. You will define and execute commissioning plans, support infrastructure bring-up, and take ownership of operations and maintenance for cutting-edge, large-scale, AI data centers. You will collaborate closely with design, construction, and hardware teams to define repeatable processes for new data center builds and lead hands-on operations to uphold the performance and reliability of our deployed infrastructure. Key Responsibilities Define and execute sequences of operations, commissioning steps, and bring-up processes for mission-critical data center facilities. Interface with the design and hardware teams to define deployment procedures tailored to each data center and hardware configuration. Oversee installation, commissioning, and operational readiness of large-scale data center campuses. Manage monitoring, maintenance, and quality control of the data center infrastructure, including high-performance liquid cooling systems. Develop on-site operations staffing strategy. Develop and enforce procedures for planed and unplanned downtime and SLAs for critical facility components. Qualifications Have 10+ years of experience in large-scale data center facility operations, commissioning, or critical infrastructure engineering. Are deeply familiar with liquid-cooled IT systems, including CDU and in-rack/in-row manifold design, bring-up, and servicing. Enjoy defining and improving infrastructure processes at the intersection of construction, hardware, and operations. Are comfortable owning infrastructure from deployment to long-term maintenance and failure recovery. Have experience responding to field issues and managing reliability through well-defined operational processes. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic. For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement . Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations. To notify OpenAI that you believe this job posting is non-compliant, please submit a report through this form . No response will be provided to inquiries unrelated to job posting compliance. We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link . OpenAI Global Applicant Privacy Policy At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.