Embedded AI Engineer, On-Device Models
Deepgram · USA | Remote
Deepgram is hiring a Embedded AI Engineer, On-Device Models in USA | Remote. The role focuses on operating systems & kernel, embedded & firmware, compilers & runtimes, databases & storage, graphics, gpu & compute, low-latency & performance.
Key responsibilities
- Write and optimize performance-critical runtime code (C, C++, and/or Rust) for embedded environments, including bare-metal and real-time operating systems such as FreeRTOS and Zephyr.
- You'll work across the stack: optimizing and compiling models for on-device inference, writing performance-critical runtime code, and squeezing every last millisecond and milliwatt out of a wide range of mobile application processors, embedded SoCs, microcontrollers, and dedicated AI accelerators.
- Build the on-device runtime plumbing: model packaging, deployment pipelines, over-the-air update mechanisms, and lightweight telemetry for devices operating with limited or intermittent connectivity.
Source and corrections
NearMetal classified and summarized this role from an official company source.
Open official source