Software Engineer, Productivity - Inference Runtime
OpenAI · San Francisco
OpenAI is hiring a Software Engineer, Productivity - Inference Runtime in San Francisco. The role focuses on compilers & runtimes, databases & storage, graphics, gpu & compute.
Key responsibilities
- C++ experience is helpful, especially for working near inference engine code, CI build issues, or performance-sensitive systems, but it is not required.
- You’ll work on the tooling and operational foundations that support model launches, inference optimizations, cloud provider integrations, and large-scale deployments across a rapidly evolving inference stack.
- You’ll also work on improving observability, rollout safety, release automation, and developer self-service tooling across a rapidly evolving inference stack.
Source and corrections
NearMetal classified and summarized this role from an official company source.
Open official source