DBS·0177
Inference Engineer
Summary
Cartesia is hiring a Inference Engineer in *HQ - San Francisco, CA. The role focuses on databases & storage, graphics, gpu & compute, low-latency & performance.
Responsibilities as stated
- Design and build low latency, scalable, and reliable model inference and serving stack for our cutting edge foundation models using Transformers, SSMs and hybrid models.
- Design and build robust inference infrastructure and monitoring for our products.
- Experience building large-scale distributed systems with high demands on performance, reliability, and observability.
Location, employment, and compensation
- Location. *HQ - San Francisco, CA — on-site
- Employment. unknown
- Compensation. Not stated on the source. NearMetal does not estimate compensation.
- How to apply. On the company’s own posting. NearMetal does not accept applications.
Why this is filed as systems software
- Design and build low latency, scalable, and reliable model inference and serving stack for our cutting edge foundation models using Transformers, SSMs and hybrid models.
- Design and build robust inference infrastructure and monitoring for our products.
- Experience building large-scale distributed systems with high demands on performance, reliability, and observability.
Source and corrections
This summary was written from the company’s own posting. NearMetal does not reproduce the full description and does not accept applications.