DBS·0056

Software Engineer, Model Inference

OpenAI · San Francisco

Databases & Query EnginesPerformance & ObservabilityStorage & Filesystems

Summary

OpenAI is hiring a Software Engineer, Model Inference in San Francisco. The role focuses on databases & storage, graphics, gpu & compute, low-latency & performance.

Responsibilities as stated

  • We are looking for an engineer who wants to take the world's largest and most capable AI models and optimize them for use in a high-volume, low-latency, and high-availability production and research environment.
  • Introduce new techniques, tools, and architecture that improve the performance, latency, throughput, and efficiency of our model inference stack.
  • Optimize our code and fleet of Azure VMs to utilize every FLOP and every GB of GPU RAM of our hardware.

Location, employment, and compensation

  • Location. San Francisco
  • Employment. unknown
  • Compensation. Not stated on the source. NearMetal does not estimate compensation.
  • How to apply. On the company’s own posting. NearMetal does not accept applications.

Why this is filed as systems software

  • We are looking for an engineer who wants to take the world's largest and most capable AI models and optimize them for use in a high-volume, low-latency, and high-availability production and research environment.
  • Introduce new techniques, tools, and architecture that improve the performance, latency, throughput, and efficiency of our model inference stack.
  • Optimize our code and fleet of Azure VMs to utilize every FLOP and every GB of GPU RAM of our hardware.

Source and corrections

This summary was written from the company’s own posting. NearMetal does not reproduce the full description and does not accept applications.

Open the company posting ↗

LAST UPDATED 3h263 roles80 companiesnext check 9h263 roles · updated 3hindex status ↗