← All jobs

Modal

Forward Deployed Engineer, ML

$180K–$250K + equity AI infraML performanceGPU

New York / San Francisco (on-site; Stockholm option) · Posted July 14, 2026

Why it's interesting

$300M+ ARR, 5x growth, a $4.65B Series C — and the FDE team ships open source (they contribute to SGLang) while tuning inference for Suno, Lovable, and Cognition. Only 2 years' ML asked; the job is GPU-level performance work at frontier labs, not slideware.

The role

Partner with leading AI companies and foundation-model labs to optimize their most demanding workloads (LLM serving, SFT/RLHF training, audio pipelines) on Modal's infrastructure. Requires 2+ years ML engineering with depth in inference optimization, GPU programming, or serving/training toolchains (vLLM, SGLang, TRL). In-person in NYC, SF, or Stockholm. Sourced from Plank FDE census (CC BY 4.0); verified on Modal's Ashby board.

Apply at Modal ↗

Applications go directly to the company. We curate; we don't broker.

Free forever. Unsubscribe anytime.