About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience.
The Role: We are looking for strong engineers with experience training production machine learning models. If you are interested in contributing to open-source projects and evolving Modal's infrastructure to train the next generation of language models, we'd love to hear from you!
Requirements: 5+ years of experience writing high-quality, high-performance code. Experience working with torch and high-level training frameworks (Huggingface, verl, slime) Experience with ML training optimization (tell us a story about eliminating data loading bottlenecks, overlapping communications with compute, rewriting a trainer to handle off-policy rollouts, etc.) Nice-to-have: familiarity with low-level operating system foundations (Linux kernel, file systems, containers, etc). Ability to work in-person, in our NYC or San Francisco office. Ability to participate in on-call rotation and respond to production incidents.
Modal is a serverless cloud infrastructure platform designed for running artificial intelligence, machine learning, and data processing workloads using code-based configuration. The company provides an environment where developers deploy functions for inference, training, batch processing, and data pipelines without managing servers or manual infrastructure setup. It offers GPU access and CPU resources that automatically scale based on workload demand, supporting applications such as large language model inference, model fine-tuning, and high-throughput computation tasks. Modal includes sandbox environments that execute isolated code for testing, evaluation, and agent-based workflows. The platform also integrates observability features that allow monitoring of logs, execution states, and performance metrics across workloads. It supports multi-cloud GPU scheduling and dynamic resource allocation to optimize compute availability across regions. Modal is used by teams building AI applications that require flexible compute scaling, low-latency execution, and simplified deployment workflows.
Salary
$150,000 - $350,000
Location
San Francisco, New York
Total raised
$111M
Last stage
Growth
Investors
Akshat Bubna
CTO and Co-Founder
No applications, no recruiter spam. Just the intro.
A few questions to make sure this role is the right shape for you. Two minutes.
I write the intro, send it to the founder, and handle the back-and-forth.
If they’re a yes, I book the chat. You show up — that’s the whole job-hunt.