DBS·0056
Software Engineer, Model Inference
Summary
OpenAI is hiring a Software Engineer, Model Inference in San Francisco. The role focuses on databases & storage, graphics, gpu & compute, low-latency & performance.
Responsibilities as stated
- We are looking for an engineer who wants to take the world's largest and most capable AI models and optimize them for use in a high-volume, low-latency, and high-availability production and research environment.
- Introduce new techniques, tools, and architecture that improve the performance, latency, throughput, and efficiency of our model inference stack.
- Optimize our code and fleet of Azure VMs to utilize every FLOP and every GB of GPU RAM of our hardware.
Location, employment, and compensation
- Location. San Francisco
- Employment. unknown
- Compensation. Not stated on the source. NearMetal does not estimate compensation.
- How to apply. On the company’s own posting. NearMetal does not accept applications.
Why this is filed as systems software
- We are looking for an engineer who wants to take the world's largest and most capable AI models and optimize them for use in a high-volume, low-latency, and high-availability production and research environment.
- Introduce new techniques, tools, and architecture that improve the performance, latency, throughput, and efficiency of our model inference stack.
- Optimize our code and fleet of Azure VMs to utilize every FLOP and every GB of GPU RAM of our hardware.
Source and corrections
This summary was written from the company’s own posting. NearMetal does not reproduce the full description and does not accept applications.