Industrial Compute

OpenAI OpenAI · AI Frontier · United States · Remote · Scaling

This role is responsible for building, scaling, and operating OpenAI's global compute infrastructure, which is critical for training frontier AI models like GPT-5.6. It involves solving complex problems across software, hardware, manufacturing, supply chain, and data center systems to improve reliability, performance, and efficiency at an extraordinary scale.

What you'd actually do

  1. Help build, scale, and operate OpenAI’s global compute infrastructure.
  2. Solve complex problems across software, hardware, manufacturing supply chain, and data center systems.
  3. Improve the reliability, performance, efficiency, and scalability of critical infrastructure.
  4. Partner with cross-functional teams to bring new compute capacity online quickly and reliably.
  5. Identify bottlenecks across technical, operational, and physical systems, and develop practical solutions.

Skills

Required

  • building, scaling, or operating complex technical systems
  • working on ambiguous, high-impact problems
  • collaborating across disciplines
  • strong technical judgment
  • bias toward execution
  • reliability
  • speed
  • safety
  • operational excellence
  • building infrastructure at unprecedented scale
  • directly support the development and deployment of frontier AI

Nice to have

  • AI infrastructure
  • high-performance computing
  • distributed systems
  • GPU clusters
  • cloud-scale platforms
  • hardware systems
  • manufacturing
  • supply chain
  • data center development
  • large capital infrastructure projects
  • civil, controls, mechanical, hardware, electrical, thermal, power, networking, or facilities engineering
  • bringing new technical platforms, data centers, factories, or large-scale systems from concept to production
  • operating in fast-moving environments

What the JD emphasized

  • frontier AI
  • compute infrastructure
  • scale

Other signals

  • scaling compute infrastructure
  • training frontier models
  • delivering compute infrastructure
  • GPU fleets
  • data center delivery