We deploy fleets of AI Infrastructure Engineers to collaborate over week-long time horizons, run 100s of experiments in parallel, and learn from them to optimize your training/ inference infrastructure to your workloads.
We've partnered with companies to speed up their GPU kernels to 6x SOTA, and improved the latency of voice models to twice the competition.
Total raised
$500K
Last stage
Pre-seed
Investors
Zayaan Mulla
Former CS student at University of Waterloo (dropped out). ML inference internship at Modular (CUDA/GPU optimization) and post-training/evals at Yutori. Competitive programming standout.
LinkedInNo applications, no recruiter spam. Just the intro.
A few questions to make sure this role is the right shape for you. Two minutes.
I write the intro, send it to the founder, and handle the back-and-forth.
If they’re a yes, I book the chat. You show up — that’s the whole job-hunt.