Compilers & RuntimesDatabases & StorageGraphics, GPU & Compute

Software Engineer, Productivity - Inference Runtime

OpenAI · San Francisco

OpenAI is hiring a Software Engineer, Productivity - Inference Runtime in San Francisco. The role focuses on compilers & runtimes, databases & storage, graphics, gpu & compute.

Key responsibilities

  • C++ experience is helpful, especially for working near inference engine code, CI build issues, or performance-sensitive systems, but it is not required.
  • You’ll work on the tooling and operational foundations that support model launches, inference optimizations, cloud provider integrations, and large-scale deployments across a rapidly evolving inference stack.
  • You’ll also work on improving observability, rollout safety, release automation, and developer self-service tooling across a rapidly evolving inference stack.

Source and corrections

NearMetal classified and summarized this role from an official company source.

Open official source