Modal
Forward Deployed Engineer, ML
Why it's interesting
$300M+ ARR, 5x growth, a $4.65B Series C — and the FDE team ships open source (they contribute to SGLang) while tuning inference for Suno, Lovable, and Cognition. Only 2 years' ML asked; the job is GPU-level performance work at frontier labs, not slideware.
The role
Partner with leading AI companies and foundation-model labs to optimize their most demanding workloads (LLM serving, SFT/RLHF training, audio pipelines) on Modal's infrastructure. Requires 2+ years ML engineering with depth in inference optimization, GPU programming, or serving/training toolchains (vLLM, SGLang, TRL). In-person in NYC, SF, or Stockholm. Sourced from Plank FDE census (CC BY 4.0); verified on Modal's Ashby board.
Applications go directly to the company. We curate; we don't broker.