Operating Systems & KernelCompilers & RuntimesDatabases & StorageVirtualization & ContainersGraphics, GPU & ComputeLow-Latency & Performance

Member of Technical Staff - ML Performance

Modal · New York, San Francisco

Modal is hiring a Member of Technical Staff - ML Performance in New York. The role focuses on operating systems & kernel, compilers & runtimes, databases & storage, virtualization & containers, graphics, gpu & compute, low-latency & performance.

Key responsibilities

  • They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale.
  • We are looking for strong engineers with experience in making ML systems performant at scale.
  • Nice-to-have: familiarity with low-level operating system foundations (Linux kernel, file systems, containers, etc).

Source and corrections

NearMetal classified and summarized this role from an official company source.

Open official source