PRF·0057
Software Engineer, Inference - Performance Optimization
Summary
OpenAI is hiring a Software Engineer, Inference - Performance Optimization in San Francisco. The role focuses on databases & storage, networking, graphics, gpu & compute, low-latency & performance.
Responsibilities as stated
- You will build cost-to-serve estimates from microbenchmarks and create tools that help cross-functional teams reason about latency, capacity, utilization, and cost tradeoffs.
- Build and refine performance models that translate microbenchmark results into cost-to-serve estimates.
- Enjoy reasoning from first principles about distributed systems, model inference, and hardware efficiency.
Location, employment, and compensation
- Location. San Francisco
- Employment. unknown
- Compensation. Not stated on the source. NearMetal does not estimate compensation.
- How to apply. On the company’s own posting. NearMetal does not accept applications.
Why this is filed as systems software
- You will build cost-to-serve estimates from microbenchmarks and create tools that help cross-functional teams reason about latency, capacity, utilization, and cost tradeoffs.
- Build and refine performance models that translate microbenchmark results into cost-to-serve estimates.
- Enjoy reasoning from first principles about distributed systems, model inference, and hardware efficiency.
Source and corrections
This summary was written from the company’s own posting. NearMetal does not reproduce the full description and does not accept applications.