NetworkingVirtualization & ContainersGraphics, GPU & ComputeLow-Latency & Performance

Backend Engineer- Inference Services

Deepgram · USA | Remote

Deepgram is hiring a Backend Engineer- Inference Services in USA | Remote. The role focuses on networking, virtualization & containers, graphics, gpu & compute, low-latency & performance.

Key responsibilities

  • Improve Deepgram’s core inference services including areas in networking, speech processing, audio transcoding, and latency and memory optimization.
  • You will design and implement secure, robust, and scalable services for speech processing; efficient, distributed compute orchestration; optimized scheduling, and more.
  • Debug complex system issues that include networking, scheduling, and high performance computing interactions.

Source and corrections

NearMetal classified and summarized this role from an official company source.

Open official source