
Explore active software engineering, data, AI, product, and design roles directly from verified employers—and make sure your resume is ready before applying.
Get an instant ATS score, missing keyword alert, and bullet rewrites tailored to your target job before submitting your application.
About the team OpenAI’s Forward Deployed Engineering team partners with customers to turn research breakthroughs into production systems. We operate at the intersection of customer delivery and core platform development. About the role Forward Deployed Engineers (FDEs) lead complex end-to-end deployments of frontier models in production alongside our most strategic customers. You will own discovery, technical scoping, system design, build, and production rollout, partnering directly with customer engineering and domain teams. You will measure success through production adoption, measurable workflow impact, and eval-driven feedback that changes product and model roadmaps. You’ll work closely with our Product, Research, Partnerships, GRC, Security, and GTM teams. This role is based in Singapore. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. 50% travel is expected. In this role you will Own technical delivery across multiple deployments from first prototype to stable production. Build full-stack systems that deliver customer value and sharpen how we learn. Embed closely with customer teams, understand their needs, and guide adoption of what you build. Scope work, sequence delivery, and remove blockers early. Make trade-offs between scope, speed, and quality; adjust plans to protect delivery. Contribute directly in the code when progress or clarity depends on it. Codify working patterns into tools, playbooks, or building blocks that others can use. Share field feedback that helps Research and Product understand where the models succeed and where they can improve. Keep teams moving through clarity and follow-through. You might thrive in this role if you Bring 5+ years of engineering or technical deployment experience that includes customer-facing work. Have scoped and delivered complex systems in fast-moving or ambiguous environments. Write and review production-grade code across frontend and backend using Python, JavaScript, or comparable stacks. Have built or deployed systems powered by LLMs or generative models and understand how model behaviour affects product experience. Simplify complexity and make fast, sound decisions under pressure. Communicate clearly with engineers, product teams, and customer stakeholders. Spot risks early and adjust without slowing down. Model calm and judgment when the stakes are high. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic. For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement . Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations. To notify OpenAI that you believe this job posting is non-compliant, please submit a report through this form . No response will be provided to inquiries unrelated to job posting compliance. We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link . OpenAI Global Applicant Privacy Policy At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.
About the Role As a Director, Compute & Infrastructure FP&A, you will own and drive the monthly forecasting process for the Compute & Infrastructure org by partnering with various stakeholders across Finance, Accounting, Tax and Engineering. You will play a critical role in planning and forecasting the company’s largest and most complex cost center ( Compute & Infrastructure ). You will collaborate cross-functionally to develop long-range infrastructure investment plans, evaluate build vs. buy decisions, and ensure capital is deployed efficiently to support rapid growth. You will also provide strategic financial guidance through scenario modeling, ROI analysis, and performance tracking, enabling leadership to make high-stakes decisions under uncertainty. What You’ll Do Own compute financial planning & Forecasting. Build and manage consolidation models for GPU/CPU capacity, storage, networking, and data center investments. Translate infrastructure roadmaps into short- and long-term financial forecasts (LRP, annual planning) Coordinate closely with Corporate FP&A on timelines and process Present insights on a monthly basis to senior management. Drive infrastructure investment decisions. Evaluate build vs. buy, vendor vs. owned infrastructure, and capacity allocation tradeoffs. Develop frameworks for investment trade-offs to guide executive decision making. Build scalable tooling & reporting. Implement stakeholder-facing dashboards to track compute spend, utilization, and efficiency metrics. Improve visibility into unit economics (e.g., cost per training run, cost per inference, cost per customer). Drive forecasting accuracy & accountability. Lead budget vs. actual analysis for compute and infrastructure spend. Identify key cost drivers (utilization, pricing, efficiency gains) and reduce forecast variance. Support close & financial reporting. Partner with Accounting to ensure accurate classification of infrastructure spend (OpEx vs CapEx). Translate complex infrastructure costs into clear insights for leadership. Enable strategic decision-making. Build scenario models to support leadership decisions on capacity scaling, new model launches, and infrastructure investments. Lead ad hoc analyses on emerging topics. You Might Thrive in This Role If You Have 10+ years in strategic finance, with experience in infrastructure, cloud, hardware, or compute-intensive environments 2+ years in investment banking Must have experience running an FP&A team at the corporate level or business unit level with significant scale. Strong financial modeling skills, particularly in capacity planning, unit economics, and scenario analysis under uncertainty. Experience supporting large-scale infrastructure or cloud spend (e.g., AWS/GCP/Azure, GPUs, data centers). Ability to translate technical concepts (compute usage, model training/inference, system architecture) into financial insights. Proficiency in Excel/Sheets, SQL, and BI tools (e.g., Tableau); experience with planning systems like Anaplan is a plus. Strong cross-functional partnership skills, especially with Engineering, Product, and Supply Chain. Familiarity with AI/ML infrastructure cost drivers and the economics of training and serving models. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic. For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement . Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations. To notify OpenAI that you believe this job posting is non-compliant, please submit a report through this form . No response will be provided to inquiries unrelated to job posting compliance. We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link . OpenAI Global Applicant Privacy Policy At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.
About the Team ChatGPT is a rapidly evolving system: new capabilities ship continuously, product surfaces change quickly, and usage patterns shift week-to-week. Supporting that pace requires infrastructure that can handle real production constraints—high concurrency, unpredictable traffic patterns, complex dependency graphs, and frequent change. The ChatGPT Infrastructure team builds and operates the platforms that enable fast iteration without compromising performance or reliability. We design shared systems, data paths, rollout mechanisms, and reliability guardrails that teams rely on to ship changes to ChatGPT at scale. We focus on high-leverage infrastructure: primitives and “golden paths” that incorporate operational lessons as defaults, so engineers don’t need to rediscover failure modes, latency pitfalls, or integration issues each time they build something new. About the Role We’re hiring Senior and Staff Engineers to design and build infrastructure systems that underlie ChatGPT and multiply the effectiveness of teams building user experiences. This is not a support-only role. It’s a platform-building role: you’ll define interfaces, develop core abstractions, and create tooling to make safe, fast iteration the norm. Your work will reduce friction, prevent regressions, improve performance, and ensure systems scale gracefully as the product grows. Where You Can Have Impact You might work on one or more of the following areas (without being restricted to any single area): Platform foundations & frameworks: Core libraries, service frameworks, and shared components that standardize system building, integration, and evolution. Scalability & performance primitives: Patterns and infrastructure that reduce tail latency, improve throughput, and keep costs predictable as demand increases. Reliability guardrails: Mechanisms that prevent outages by design—rate limiting, load shedding, dependency isolation, backpressure, safe fallbacks, and robust regression controls. Developer productivity via golden paths: Paved roads for common workflows (data access patterns, service integration, request lifecycle) that are fast, safe, and easy to use. Observability & debugging systems: Instrumentation, metrics models, and investigative tooling that turn vague symptoms (“it’s slow”) into precise, actionable diagnoses. Safe change management: Deployment and rollout systems that support rapid iteration with confidence—progressive delivery, automated verification, and fast rollback strategies. Interface and contract design across boundaries: Clean APIs and stable contracts that reduce coupling and enable independent evolution across a complex ecosystem. What You’ll Do Build and evolve infrastructure platforms used by many engineers and services. Translate real-world constraints into clean abstractions: simple APIs, enforceable contracts, safe defaults. Drive improvements in reliability and performance through principled design, measurement, and iterative hardening. Partner with engineering and product teams to identify systemic pain points and develop reusable solutions. Own outcomes end-to-end: design → implementation → rollout → operational maturity. Qualifications Minimum Qualifications Experience building and operating large-scale distributed systems in production (high throughput, concurrency, and failure handling). Strong fundamentals in systems design, including caching, consistency, queueing/backpressure, and resilient dependency management. Ability to reason about performance (latency distributions, tail behavior, bottlenecks) and translate analysis into concrete engineering work. Track record of building platforms or shared infrastructure that improves velocity and correctness for other teams. Excellent communication and collaboration skills—aligning on interfaces, navigating tradeoffs, and driving cross-team execution. Preferred Qualifications Experience designing paved roads / golden paths (frameworks, libraries, self-serve tooling) that shape engineering behavior at scale. Deep understanding of reliability techniques: graceful degradation, circuit breakers, load shedding, rate limiting, and fault isolation. Experience building systems for safe iteration: progressive delivery, correctness checks, automated rollout gates, and production validation. Strong instincts for API and contract design—how to create interfaces that are stable, evolvable, and hard to misuse. Prior work that demonstrates “force multiplier” impact: enabling many teams via a small set of well-crafted primitives. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic. For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement . Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations. To notify OpenAI that you believe this job posting is non-compliant, please submit a report through this form . No response will be provided to inquiries unrelated to job posting compliance. We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link . OpenAI Global Applicant Privacy Policy At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.
About the Team OpenAI is building the infrastructure foundation for the next generation of AI. The Data Center Engineering team defines the strategy, reference architectures, technical requirements, and delivery standards for the large-scale data centers that support OpenAI research, products, and infrastructure partners. As a Data Center Controls Network Engineer, you will design, validate, and scale the controls and OT network architectures that support high-density AI data centers. You will work across controls systems, OT infrastructure, telemetry, commissioning, deployment, and operations, partnering with mechanical, electrical, IT/networking, security, and external delivery teams. About the Role We are seeking a mid to senior OT Network Engineer with a strong controls systems background to lead the design and operation of resilient, secure, and scalable OT network architectures for high-density AI data centers. This role translates compute, power, cooling, and operational requirements into practical OT network designs, evaluates vendor solutions, and drives technical decisions across controls infrastructure, telemetry, commissioning, and operations. The ideal candidate has strong hands-on experience in mission-critical OT environments, including industrial networking, virtualized infrastructure, and OT network operations, with expertise in routing, switching, segmentation, firewall policy, time synchronization, monitoring, and network lifecycle support. Key Responsibilities Define controls, automation, and OT network requirements for AI data center campuses. Develop reference architectures, engineering standards, and reusable design templates. Review and develop basis-of-design and functional design documents, including OT network diagrams, IP/VLAN schemes, telemetry architectures, data flow diagrams, and commissioning requirements. Design OT and infrastructure network architectures, including physical topology, logical topology, IP addressing, subnetting, VLANs, routing, switching, redundancy, segmentation, firewall policy coordination, out-of-band management, monitoring, and remote access patterns. Develop day-two network operations requirements, including change management, configuration backups, golden configurations, monitoring thresholds, firmware lifecycle, rollback plans, and post-change validation. Partner with electrical, mechanical, IT/networking, security, and operations teams to ensure OT network systems align with GPU deployments, campus-wide telemetry, and failure-domain isolation requirements. Define integration patterns and protocol requirements across BACnet/IP, BACnet MSTP, Modbus TCP/RTU, OPC UA, IEC-61850 MMS/GOOSE, MQTT, SNMP, syslog, NTP/PTP, IRIG-B, and vendor-specific interfaces. Lead technical evaluation of controls integrators, network equipment suppliers, design consultants, contractors, and commissioning agents Review network equipment submittals, configurations, firmware assumptions, certifications, test reports, and quality documentation. Support factory witnessed testing (FWT), site acceptance testing, network readiness checks, failover testing, and integrated systems testing. Troubleshoot complex controls network issues including packet loss, latency, duplicate IPs, routing errors, firewall drops, protocol incompatibilities, time synchronization drift, and intermittent device communication failures. Qualifications 8+ years of relevant experience in controls engineering, industrial automation, OT networking, mission-critical facilities, or similar critical infrastructure environments. Strong expertise in resilient OT network architecture, implementation, troubleshooting, and lifecycle support. Experience with OT/IT boundary design, secure enterprise integration, firewall policy design, redundant topologies, out-of-band management, and monitoring. Hands-on experience with Layer 3 OT network design, including IP addressing, subnetting, routing, VRFs, ACLs, inter-VLAN traffic control, and network segmentation. Hands-on experience with Layer 2 security and switching controls, including MACsec, port security, loop prevention, and switch-level access control. Hands-on experience in designing resilient OT network topologies using industrial redundancy protocols and architectures such as PRP, HSR, Cisco REP, RSTP/MSTP, and ring or star topologies. Hands-on experience in designing resilient infrastructure network architectures using HSRP/VRRP, spine-leaf topologies, redundant uplinks, and failure-domain isolation. Hands-on experience with industrial and infrastructure network equipment such as Cisco switches/routers, Juniper switches/routers, Palo Alto firewalls, Rockwell Automation Stratix switches, Siemens Ruggedcom or comparable industrial networking platforms. Experience with network management and observability platforms such as Cisco Catalyst Center (DNA Center), Palo Alto Panorama, Juniper Mist, industrial NMS tools, packet brokers, and OT monitoring platforms. Hands-on experience with industrial Ethernet, VPN tunneling, IPsec-based connectivity, and secure remote access. Hands-on experience with virtualized OT or controls server environments such as VMware vSAN, Microsoft Azure Stack HCI / Hyper-V, or comparable infrastructure platforms. Experience with industrial communication and OT infrastructure protocols, including BACnet/IP, BACnet MSTP, Modbus TCP/RTU, OPC UA, IEC-61850 MMS/GOOSE, MQTT, SNMP, syslog, NTP/PTP, IRIG-B, and vendor-specific interfaces, and strong understanding of their behavior across OT network architectures. Experience reviewing and producing technical design documentation, commissioning plans, and acceptance test procedures. Experience with factory witnessed testing, site acceptance testing, failover testing, telemetry validation, protocol compatibility testing, and root-cause analysis. Ability to use logs, packet captures, and field observations to make sound technical decisions and communicate risk clearly. Bachelor’s degree in Electrical Engineering, Computer Engineering, Network Engineering, Systems Engineering, or a related discipline. Preferred Skills Master's degree in Electrical Engineering, Computer Engineering, Network Engineering, Systems Engineering, or a related discipline. Experience leading multi-campus OT network integration, commissioning, and operations across cross-functional teams, contractors, vendors, and delivery partners. Relevant networking certifications such as Cisco CCNA/CCNP, Palo Alto PCNSA/PCNSE, Juniper JNCIA/JNCIS, or similar networking credentials. Cybersecurity certifications such as CISSP, GICSP, ISA/IEC 62443, CompTIA Security+, or similar cybersecurity credentials are a plus. Experience with network automation, Git-based configuration management, and Infrastructure as Code (IaC) using tools such as Ansible, Terraform, Python, or similar to support scalable OT network deployment and lifecycle management. Experience with scripting, APIs, and automation workflows that improve OT network operations. Experience using AI agents or MCP-connected tools to support telemetry analysis,and troubleshooting. Experience with relational database systems such as PostgreSQL, SQL Server, MySQL, or similar platforms used for OT telemetry, historian integrations, troubleshooting, and reporting. Work Environment and Travel This role requires periodic travel to data center campuses, vendors, labs, construction sites, commissioning activities, and controls/network cutovers. The engineer should be comfortable working across office, lab, construction, and live data center environments, including PPE, lockout/tagout, cyber hygiene, and change-control requirements. Work may include time-sensitive support during commissioning, startup, vendor testing, cutovers, network changes, telemetry issues, automation failures, and operational events. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic. For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement . Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations. To notify OpenAI that you believe this job posting is non-compliant, please submit a report through this form . No response will be provided to inquiries unrelated to job posting compliance. We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link . OpenAI Global Applicant Privacy Policy At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.
About the Team OpenAI’s Infrastructure organization builds and evaluates the systems that power advanced AI workloads. We work closely with hardware, modeling, and architecture teams to ensure that new platforms deliver real-world performance aligned with workload needs. Our team focuses on understanding workload behavior across evolving hardware platforms—bridging the gap between theoretical capability and observed system performance. About the Role We are seeking a Workload Porting & Performance Engineer to evaluate new hardware platforms by porting benchmarks and real-world workloads, analyzing performance, and identifying system bottlenecks. In this role, you will bring up workloads on new systems, characterize performance behavior, and adapt workloads to better utilize hardware capabilities. You will play a critical role in validating new platforms and ensuring that performance aligns with expectations across compute, memory, and networking subsystems. This role requires strong hands-on experience with performance analysis, workload optimization, and system-level debugging across hardware and software boundaries. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance. Key Responsibilities Port and enable benchmarks and real-world workloads on new hardware platforms. Evaluate system performance across compute, memory, storage, and networking subsystems. Identify and analyze performance bottlenecks and inefficiencies. Adapt and optimize workloads to better utilize hardware capabilities. Develop and run performance experiments and profiling workflows. Compare expected vs. observed performance and provide feedback to: hardware architecture teams performance modeling teams system and software engineers. Debug issues across the stack, including software, runtime, and hardware interactions. Provide actionable insights to guide platform readiness and deployment decisions. Qualifications Experience with performance analysis, benchmarking, or workload optimization. Strong understanding of system architecture, including CPU/GPU, memory, and I/O subsystems. Experience porting or adapting workloads across different hardware platforms. Familiarity with profiling tools and performance debugging techniques. Ability to identify root causes of performance issues across hardware/software boundaries. Experience working in large-scale or distributed system environments. Preferred Skills Experience with AI/ML workloads, including training or inference systems. Familiarity with GPU or accelerator-based systems. Experience working with low-level performance tools (profilers, tracing, microbenchmarks). Background in systems software, compilers, or runtime optimization. Experience collaborating with hardware and architecture teams on performance validation. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic. For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement . Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations. To notify OpenAI that you believe this job posting is non-compliant, please submit a report through this form . No response will be provided to inquiries unrelated to job posting compliance. We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link . OpenAI Global Applicant Privacy Policy At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.
About the Team OpenAI’s Hardware organization develops system and infrastructure solutions tailored to the demands of advanced AI workloads. We work across the full stack—from silicon to system integration—partnering closely with internal teams and external vendors to define and deliver next-generation AI infrastructure. Our team focuses on defining scalable, high-performance system architectures and reference designs that balance performance, cost, and operational efficiency across rapidly evolving technologies. About the Role We are seeking a 3P Architect to define and drive rack- and cluster-level reference designs in collaboration with external partners. This role is responsible for translating workload requirements and system-level goals into concrete architectures, aligning partners on critical design attributes, and ensuring vendor roadmaps meet our infrastructure needs. You will work closely with performance modeling and internal architecture teams to evaluate tradeoffs, while owning the end-to-end definition and execution of third-party system designs. This includes identifying gaps in current technologies, driving vendor development, and shaping future infrastructure capabilities. This role requires strong system intuition, cross-functional leadership, and the ability to operate effectively across internal teams and external ecosystems. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance. Key Responsibilities Define rack- and cluster-level reference architectures for AI infrastructure deployments. Translate workload requirements into clear system design specifications and partner deliverables. Collaborate with performance modeling teams to evaluate architectural tradeoffs and system behaviors. Align internal stakeholders and external partners on critical system attributes (performance, cost, power, reliability, scalability). Identify gaps in current technology offerings and drive vendors (ODM/JDM, silicon, networking) to close those gaps. Influence and shape vendor roadmaps to meet future infrastructure needs. Track emerging technologies and evaluate their applicability to AI systems. Define and lead proof-of-concept (PoC) efforts to validate new architectures and technologies. Act as a key interface between OpenAI and external partners, ensuring execution against design intent. Qualifications Have strong experience in system architecture for large-scale infrastructure or data center environments. Understand AI workload characteristics and how they map to system-level design decisions. Are comfortable working with performance modeling outputs to inform architectural direction. Have experience working with or managing hardware vendors (ODM/JDM, silicon, networking). Can drive alignment across multiple stakeholders with competing constraints. Have a track record of turning ambiguous requirements into clear, executable system designs. Are proactive in identifying gaps and driving solutions across organizational boundaries. Preferred Skills Experience defining rack- or cluster-level systems for hyperscale or AI workloads. Familiarity with accelerators (GPUs/ASICs), interconnects, and data center networking architectures. Experience influencing vendor roadmaps and reference designs. Background in infrastructure deployment, hardware engineering, or systems integration. Experience leading PoCs or early-stage hardware validation efforts. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic. For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement . Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations. To notify OpenAI that you believe this job posting is non-compliant, please submit a report through this form . No response will be provided to inquiries unrelated to job posting compliance. We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link . OpenAI Global Applicant Privacy Policy At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.
About the Team OpenAI is building the infrastructure foundation for the next generation of AI. The Data Center Engineering team defines the strategy, reference architectures, technical requirements, and delivery standards for the large-scale data centers that support OpenAI research, products, and infrastructure partners. As a Data Center Infrastructure Engineering Program Manager, you will help turn complex infrastructure strategy into executable programs across electrical, mechanical, controls, network, hardware, construction, commissioning, deployment, and operations workstreams. You will partner with research, hardware engineering, data center engineering, site development, supply chain, security, EHS, finance, legal, operations, and external delivery partners to bring OpenAI's infrastructure vision to life. About the Role We are looking for an Engineering Program Manager (EPM) to lead assigned infrastructure programs focused on production and non-production network integration, controls coordination, and the design and deployment of data hall or whitespace facilities. The EPM will support functional Directly Responsible Individuals (DRIs) across network, controls, structural, electrical, and mechanical disciplines. Key responsibilities include coordinating assigned workstreams and program controls, maintaining risks and interfaces, and supporting readiness within the network and data hall deployment track. The ideal candidate thrives on bringing structure to complex environments characterized by ambiguous technical requirements, large partner ecosystems, tight deadlines, and high operational stakes. This individual must be adept at keeping teams aligned on decisions, risks, dependencies, schedules, and readiness criteria, and escalating gaps or decision points when needed. Candidates should have a proven track record of managing technically challenging engineering programs across major lifecycle phases, including design, validation, procurement, construction, commissioning, deployment, and operational handoff. Key Responsibilities Translate assigned infrastructure goals into clear workstream charters, scopes, milestones, owners, decision points, success metrics, resourcing assumptions, and execution plans. Build and maintain integrated execution plans for assigned programs covering network, controls, data hall design, whitespace deployment, commissioning preparation, and deployment readiness. Support coordination across network, controls, structural, electrical, mechanical, hardware integration, construction, commissioning, and operations teams. Work with third-party design teams to define design milestones from concept through detailed design, including basis-of-design development, requirements tracking, design reviews, technical comment resolution, change management, and release readiness. Maintain the dependency map, issue log, risk register, action tracker, and decision log for assigned network and data hall workstreams. Coordinate design and review milestones for network rooms, non-production network services, OT / IT interface points, rack deployment assumptions, telemetry interfaces, controls dependencies, and data hall deployment packages. Track building-level network and support-space interfaces such as MPOE, MMR, Network Core, WAN, support rooms, and associated handoff points where they affect assigned programs. Support the network and controls DRIs by organizing reviews, resolving cross-discipline gaps, surfacing decisions, and keeping partner deliverables aligned to schedule. Manage partner and vendor deliverables such as submittals, interface packages, installation assumptions, turn-up plans, readiness evidence, field issue logs, and corrective action tracking. Drive readiness tracking for assigned 1P, 3P, colo, and selected CSP programs, including bring-up sequencing, installation readiness, access dependencies, maintenance windows, and first-use criteria. Prepare clear status updates, dashboards, and executive-ready summaries for the Industrial Compute lead and project stakeholders. Capture lessons learned from deployment and handoff activities and feed them back into playbooks, standards, and interface definitions. Qualifications Extensive experience in engineering program management, technical program management, mission-critical infrastructure delivery, data center deployment, or comparable complex execution environments, typically gained through 10+ years of relevant work or equivalent depth of experience. Proven ability to operate within ambiguous, cross-functional engineering programs with shifting requirements, urgent timelines, and high-stakes operational risk, driving from concept through design, validation, procurement, construction, commissioning, deployment, and operations. Proven experience coordinating cross-functional programs that include network, controls, mechanical, electrical, structural, construction, commissioning, or operations participants. Strong technical fluency in at least several of the following areas: data hall deployment, non-production network, production network interfaces, controls coordination, telemetry, rack deployment, mission-critical support spaces, and infrastructure handoff. Experience building and maintaining integrated schedules, dependency maps, risk registers, decision logs, readiness trackers, and partner action plans. Experience coordinating external partners, vendors, design firms, delivery teams, or operators in a multi-party infrastructure environment. Ability to understand complex technical tradeoffs, ask strong questions, identify hidden dependencies, and help teams move toward clear decisions without needing to be the sole technical owner. Excellent written and verbal communication skills, with the ability to produce crisp status reporting and drive action across matrixed teams. Bachelor's degree in Engineering, Computer Science, Construction Management, Operations, Business, or a related technical or quantitative field, or equivalent practical experience. Preferred Skills Direct experience with hyperscale data centers, AI infrastructure, HPC environments, colocation, or partner-delivered data center programs. Experience supporting high-density data hall or HAC-related design and deployment efforts. Experience with non-production network and support-service readiness for large-scale infrastructure deployments. Experience coordinating controls integration, network-room layouts, telemetry interfaces, rack deployment packages, or early operational handoff. Comfort with technical documentation such as one-line diagrams, P&IDs, controls sequences, network diagrams, equipment specifications, interface control documents, telemetry schemas, test procedures, and commissioning scripts. Familiarity with 1P, 3P, colocation, and cloud service provider delivery models and the differing owner-partner interface expectations in each. Work Environment and Travel This role may require periodic travel to data center campuses, manufacturing partners, equipment suppliers, laboratories, construction sites, commissioning activities, and partner program reviews. The program manager should be comfortable working across office, lab, manufacturing, construction, and operating data center environments, including environments that require PPE, safety briefings, change-control discipline, and coordination with site operations. Work may include time-sensitive escalations during design reviews, procurement, manufacturing validation, commissioning, startup, production deployment, vendor testing, operational readiness, or operational incidents. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic. For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement . Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations. To notify OpenAI that you believe this job posting is non-compliant, please submit a report through this form . No response will be provided to inquiries unrelated to job posting compliance. We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link . OpenAI Global Applicant Privacy Policy At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.
About the Team OpenAI’s Hardware organization develops system and infrastructure solutions designed for the unique demands of advanced AI workloads. We work closely with architecture, infrastructure, and vendor teams to evaluate system performance and guide critical design decisions. Our team focuses on building and applying performance modeling frameworks to understand system behavior, quantify tradeoffs, and support next-generation infrastructure design. About the Role We are seeking an Performance Modeling Engineer to support the development and application of modeling tools used to evaluate AI system performance and inform architectural decisions. In this role, you will partner closely with Senior Performance Modeling Engineers and the Performance Modeling Lead to analyze system behavior, run simulations and analytical models, and help evaluate tradeoffs across compute, memory, networking, and storage. You will contribute to building modeling frameworks while developing a strong foundation in system architecture and AI infrastructure. This role is ideal for early-career engineers with 1–2 years of experience in software engineering, systems analysis, or performance modeling who are excited to grow in large-scale infrastructure and hardware/software systems. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance. Key Responsibilities Support the development and maintenance of performance modeling tools and frameworks Assist in building models to evaluate system behavior across compute, memory, networking, and interconnect subsystems Help analyze distributed system scaling behavior and identify performance bottlenecks Run simulations and analytical models to support architecture and infrastructure decisions Partner with senior engineers to evaluate design tradeoffs across hardware and system components Interpret modeling outputs and help translate findings into clear recommendations Validate models using benchmarking data and real system performance measurements Improve modeling workflows, documentation, and usability for broader team adoption Collaborate cross-functionally with hardware, infrastructure, and architecture teams Continuously build technical depth across AI infrastructure, system architecture, and performance analysis Qualifications 1–2 years of experience in software engineering, systems modeling, performance analysis, or related technical work Strong programming skills and experience building technical tools, scripts, or frameworks Familiarity with system architecture fundamentals such as compute, memory, and networking Ability to reason about system performance, bottlenecks, and scaling behavior Strong analytical and problem-solving skills with comfort working in quantitative environments Ability to learn quickly and work effectively across technical teams Preferred Skills Exposure to AI/ML workloads, distributed systems, or large-scale infrastructure Experience with simulation tools, benchmarking, profiling, or performance analysis Familiarity with data center systems, server architecture, or hardware platforms Interest in system architecture and hardware/software co-design Internship or early professional experience in performance engineering, infrastructure, or systems design About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic. For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement . Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations. To notify OpenAI that you believe this job posting is non-compliant, please submit a report through this form . No response will be provided to inquiries unrelated to job posting compliance. We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link . OpenAI Global Applicant Privacy Policy At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.
About the Team OpenAI’s Hardware organization develops system and infrastructure solutions designed for the unique demands of advanced AI workloads. We work closely with architecture, infrastructure, and vendor teams to evaluate system performance and guide critical design decisions. Our team focuses on building and applying performance modeling frameworks to understand system behavior, quantify tradeoffs, and inform next-generation infrastructure design. About the Role We are seeking Performance Modeling Engineers to develop and apply modeling tools that evaluate AI system performance and inform architectural decisions. In this role, you will work closely with the Performance Modeling Lead and partner teams to analyze system behavior, run simulations or analytical models, and help quantify tradeoffs across compute, memory, networking, and storage. You will contribute to building modeling frameworks and applying them to real-world questions that impact system design and vendor decisions. This role is well-suited for engineers with strong software or modeling backgrounds who are interested in developing deeper expertise in system architecture and AI infrastructure. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance. Key Responsibilities Develop and maintain performance modeling tools and frameworks. Build models to evaluate system behavior across: compute, memory, and interconnect subsystems distributed system scaling and bottlenecks. Run simulations and analytical models to support architectural tradeoff analysis. Collaborate with performance modeling lead and system architects to answer forward-looking design questions. Analyze and interpret modeling outputs, translating results into actionable insights. Validate models against real system measurements and workload behavior. Contribute to improving modeling fidelity, usability, and scalability. Qualifications Strong software engineering or modeling background (e.g., simulation, systems modeling, or performance analysis). Familiarity with system architecture fundamentals (compute, memory, networking). Experience with programming and building technical tools or frameworks. Ability to reason about performance bottlenecks and scaling behavior. Strong analytical skills and comfort working with quantitative models. Ability to collaborate across teams and learn new system domains quickly. Preferred Skills Exposure to AI/ML workloads or distributed systems. Experience with simulation tools, performance modeling, or systems analysis. Familiarity with data center infrastructure or large-scale systems. Experience working with performance data, benchmarking, or profiling tools. Interest in system architecture and hardware/software co-design. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic. For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement . Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations. To notify OpenAI that you believe this job posting is non-compliant, please submit a report through this form . No response will be provided to inquiries unrelated to job posting compliance. We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link . OpenAI Global Applicant Privacy Policy At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.
About the Team OpenAI’s Hardware organization develops system and infrastructure solutions designed for the unique demands of advanced AI workloads. We work closely with research, software, and external hardware partners to shape the next generation of AI systems, from silicon through full-scale deployments. Our team focuses on understanding and optimizing performance across the full system stack—ensuring that architectural decisions are grounded in rigorous, quantitative analysis of real-world workloads. About the Role We are seeking a Performance Modeling Lead to build and lead a small, high-impact team responsible for answering forward-looking architectural questions across AI infrastructure systems. You will develop modeling frameworks and methodologies to evaluate system-level tradeoffs and guide key design decisions. Your work will directly influence reference architectures, vendor designs, and long-term infrastructure strategy. This role sits at the intersection of AI workloads, system architecture, and quantitative modeling, and requires strong technical judgment, ownership, and the ability to translate complex analysis into clear, actionable guidance. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance. Key Responsibilities Build and own a performance modeling framework/toolchain to evaluate AI systems across multiple levels of abstraction. Analyze and quantify architectural tradeoffs across compute, memory, networking, storage, and system topology. Develop performance models to guide decisions on: scale-up vs. scale-out architectures interconnect and network design memory hierarchy and system balance. Translate modeling outputs into clear recommendations for internal teams and external hardware vendors. Influence reference designs and vendor roadmaps through data-driven insights. Partner closely with machine learning, systems, and hardware teams to understand workload characteristics and requirements. Lead and grow a small team (2–3 engineers), setting technical direction and maintaining high standards for modeling rigor. Continuously improve modeling fidelity by validating against real system behavior and measurements. Qualifications Have experience owning or building performance modeling frameworks used to drive real system design decisions. Have deep knowledge of AI/ML workloads, including training and/or inference at scale. Understand system-level tradeoffs across compute, memory, and networking in large-scale distributed systems. Are comfortable working across abstraction layers—from workload behavior to hardware implementation. Have experience using modeling (analytical or simulation) to inform architectural decisions. Can operate in ambiguous problem spaces and turn open-ended questions into structured analysis. Communicate clearly and influence both internal teams and external partners. Preferred Skills Experience working with hardware vendors (ODM/JDM, silicon, networking). Background in data center infrastructure or hyperscale systems. Familiarity with accelerators (GPUs/ASICs) and interconnects (e.g., NVLink, InfiniBand, Ethernet). Experience influencing hardware roadmaps or reference architectures. Prior experience leading or mentoring engineers. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic. For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement . Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations. To notify OpenAI that you believe this job posting is non-compliant, please submit a report through this form . No response will be provided to inquiries unrelated to job posting compliance. We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link . OpenAI Global Applicant Privacy Policy At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.
About the Team: OpenAI's mission is to ensure that artificial general intelligence (AGI) benefits all of humanity. Our API is the industry's most widely adopted AI platform, empowering startups, indie developers, and Fortune 500 companies alike. Through multimodal APIs—spanning real-time interactions, text-to-speech (TTS), speech generation, and image generation—we enable users to harness the full potential of diverse AI modalities effectively and at scale. About the Role: We are seeking an Engineering Manager to lead our multimodal API product suite. Your team will be responsible for delivering innovative APIs across real-time processing, speech transcription, speech generation, and image creation. You will own the product roadmap for how we evolve our multimodal API offerings, and you will build the products that allow developers to reach millions of end users through AI audio, video, and images. In this role, you will: Build, mentor, and grow a high-performing engineering team focused on multimodal API products – including our realtime API, our transcription models (Whisper), our speech generation models (TTS), and our image generation APIs (DALLE and native 4o). Collaborate closely with product managers, designers, and other stakeholders to define the strategic vision and product roadmap. Work closely with our research teams to improve our core multimodal models for API customer use cases. Guide technical and architectural decisions, emphasizing scalability, robustness, and user experience. Foster a culture of innovation, continuous improvement, and accountability within your team. Qualifications: Proven experience managing engineering teams that deliver complex, high-quality products at scale. Strong technical background and proficiency in modern software engineering practices and system architecture. Excellent collaboration and communication skills to effectively coordinate across diverse teams and stakeholders. Familiarity with or strong interest in multimodal AI, including speech technologies, real-time systems, and image generation. Ability to operate effectively in a fast-paced, ambiguous startup environment. Preferred Qualifications: Experience developing multimodal systems or APIs in AI/ML domains, especially around image generation, audio generation, or speech transcription. Familiarity with real-time streaming technologies, audio processing, and computer vision. Hands-on experience with cloud platforms and distributed architectures. Location: This role is in-person at our San Francisco office. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic. For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement . Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations. To notify OpenAI that you believe this job posting is non-compliant, please submit a report through this form . No response will be provided to inquiries unrelated to job posting compliance. We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link . OpenAI Global Applicant Privacy Policy At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.
About the Team OpenAI is building the infrastructure foundation for the next generation of AI. The Data Center Engineering team defines the strategy, reference architectures, technical requirements, and delivery standards for the large-scale data centers that support OpenAI research, products, and infrastructure partners. As a Data Center Infrastructure Electrical Engineer, you will help define, validate, and scale the electrical power systems that support high-density AI compute. You will translate evolving compute requirements into practical facility and rack-power architectures, evaluate new technologies and vendor solutions, and drive technical decisions across design, manufacturing validation, construction, commissioning, deployment, and operations. This role is best suited for a senior hands-on engineer with deep experience in mission-critical power systems, strong judgment under ambiguity, and the ability to connect facility infrastructure, hardware requirements, controls, telemetry, reliability, and operations. About the Role We are seeking a senior electrical infrastructure engineer to lead the development of reliable, scalable, and efficient power architectures for high-density, liquid-cooled AI data centers. The ideal candidate has strong practical experience with critical electrical systems at data centers or comparable industrial scale, including medium-voltage and low-voltage distribution, utility interfaces, backup power, UPS and battery systems, rack power delivery, grounding, protection, controls, and monitoring systems. You should be comfortable moving between long-range architecture, detailed engineering review, lab validation, vendor qualification, field deployment, and operational troubleshooting. Key Responsibilities Design and optimize electrical topologies and equipment strategies that reduce cost, accelerate schedules, improve efficiency, increase scalability, and maintain high reliability and maintainability. Review and develop basis-of-design documents, single-line diagrams, equipment specifications, commissioning plans, and power system studies including load flow, short circuit, protection coordination, arc flash, and grid transient compliance. Lead technical evaluation of equipment vendors, manufacturers, design consultants, commissioning agents, contractors, and test labs. Review submittals, schematics, certifications, quality records, and test reports. Evaluate AC and DC power distribution options for high-density compute, including rectifier architectures, busway, rack power shelves, power supplies, cable management, redundancy strategies, serviceability, and fault isolation. Partner with mechanical, cooling, controls, hardware, networking, construction, and operations teams to ensure electrical systems support liquid-cooled GPU rack deployments and reliable facility operation. Drive FAT, SAT, witness testing, burn-in, reliability testing, interoperability testing, and integrated systems testing for critical infrastructure and rack deployments. Help guide and operate an R&D program to validate new equipment, rack designs, telemetry, operating envelopes, failure scenarios, and facility-to-hardware interactions. Define telemetry and controls requirements including metering, waveform capture, breaker and relay status, UPS and battery health, generator status, rack power, CDU status, coolant conditions, alarms, and control states. Analyze lab data, operational incidents, power quality events, nuisance trips, thermal excursions, controls alarms, and rack-level failures to improve designs, procedures, vendor quality, and reliability models. Create engineering standards, test procedures, commissioning scripts, operating procedures, decision records, risk registers, and concise executive summaries. Provide senior technical escalation support during design reviews, construction, manufacturing validation, commissioning, energization, deployment, and operational events. Raise the technical bar across partner teams in electrical design, safety, testing discipline, documentation quality, reliability, and operational rigors Qualifications 10+ years of electrical engineering experience in mission-critical facilities such as data centers or comparable critical infrastructure. Bachelor’s degree in electrical engineering, power systems, or a related discipline. Strong hands-on experience with MV/LV distribution, transformers, switchgear, switchboards, generators, UPS systems, batteries, PDUs, RPPs, busway, PSUs, grounding, bonding, and protection systems. Experience producing and reviewing electrical designs, studies, equipment specifications, commissioning plans, test procedures, and acceptance criteria. Experience leading FWTs, FATs, SATs, integrated systems testing, failure-mode testing, and root-cause investigations. Strong understanding of codes, standards, utility coordination, safety requirements, and AHJ processes. Ability to evaluate conflicting technical tradeoffs involving reliability, cost, schedule, constructability, maintainability, and scalability. Strong written and verbal communication skills with the ability to explain complex issues clearly to technical and non-technical stakeholders. Comfortable operating in fast-moving environments with incomplete information and changing priorities. Preferred Skills 15+ years of relevant experience. Advanced degree in electrical engineering, power systems, or a related discipline. Professional Engineer license, Chartered Engineer status, or comparable professional certification. Direct experience with hyperscale data centers, GPU clusters, high-density rack deployments, or liquid-cooled compute environments. Experience qualifying new electrical products from prototype through certification, manufacturing ramp, deployment, and lifecycle support. Experience with BMS, EPMS, DCIM, SCADA, PLCs, protective relays, waveform capture, data historians, or reliability analytics platforms. Experience with modular power systems, rack-level conversion, BESS, HVDC/LVDC, or other advanced power architectures. Experience managing global, multi-site deployment programs or manufacturing quality programs. Familiarity with grid constraints, demand response, renewable integration, or low-carbon infrastructure strategies. Work Environment and Travel This role requires periodic travel to data center campuses, manufacturing partners, equipment suppliers, laboratories, construction sites, and commissioning activities. The person in this role should be comfortable working in office, lab, manufacturing, construction, and operating data center environments, including environments that require PPE, safety briefings, lockout/tagout discipline, and coordination with site operations. Work may include time-sensitive technical escalations during commissioning, energization, production deployment, vendor testing, or operational incidents. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic. For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement . Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations. To notify OpenAI that you believe this job posting is non-compliant, please submit a report through this form . No response will be provided to inquiries unrelated to job posting compliance. We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link . OpenAI Global Applicant Privacy Policy At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.