CMP·0062
Software Engineer, Productivity - Inference Runtime
Summary
OpenAI is hiring a Software Engineer, Productivity - Inference Runtime in San Francisco. The role focuses on compilers & runtimes, databases & storage, graphics, gpu & compute.
Responsibilities as stated
- C++ experience is helpful, especially for working near inference engine code, CI build issues, or performance-sensitive systems, but it is not required.
- You’ll work on the tooling and operational foundations that support model launches, inference optimizations, cloud provider integrations, and large-scale deployments across a rapidly evolving inference stack.
- You’ll also work on improving observability, rollout safety, release automation, and developer self-service tooling across a rapidly evolving inference stack.
Location, employment, and compensation
- Location. San Francisco
- Employment. unknown
- Compensation. Not stated on the source. NearMetal does not estimate compensation.
- How to apply. On the company’s own posting. NearMetal does not accept applications.
Why this is filed as systems software
- C++ experience is helpful, especially for working near inference engine code, CI build issues, or performance-sensitive systems, but it is not required.
- You’ll work on the tooling and operational foundations that support model launches, inference optimizations, cloud provider integrations, and large-scale deployments across a rapidly evolving inference stack.
- You’ll also work on improving observability, rollout safety, release automation, and developer self-service tooling across a rapidly evolving inference stack.
Source and corrections
This summary was written from the company’s own posting. NearMetal does not reproduce the full description and does not accept applications.