Software Engineer, Model Inference
OpenAI · San Francisco
OpenAI is hiring a Software Engineer, Model Inference in San Francisco. The role focuses on databases & storage, graphics, gpu & compute, low-latency & performance.
Key responsibilities
- We are looking for an engineer who wants to take the world's largest and most capable AI models and optimize them for use in a high-volume, low-latency, and high-availability production and research environment.
- Introduce new techniques, tools, and architecture that improve the performance, latency, throughput, and efficiency of our model inference stack.
- Optimize our code and fleet of Azure VMs to utilize every FLOP and every GB of GPU RAM of our hardware.
Source and corrections
NearMetal classified and summarized this role from an official company source.
Open official source