> Fuse Energy · London, United Kingdom (Remote) · — · Posted 2026-07-20
Workplace: remote
Fuse Energy is a forward-thinking renewable energy startup on a mission to deliver a terawatt of renewable energy - fast. We're combining first-principles thinking with cutting-edge technology to build a radically better energy system. We raised $210M from top-tier investors including Multicoin, Balderton, Lakestar, Accel, Creandum, Lowercarbon, Ribbit, Box Group and strategic angels like Nico Rosberg, the Co-Founder of Solana and GPs behind Meta, Revolut, Spotify, Uber and more.
As data centres become one of the largest and fastest-growing sources of electricity demand, Fuse is expanding into high-performance compute infrastructure that sits at the intersection of energy and AI. We're building the GPU/CUDA performance layer and the inference serving layer at the same time, from scratch - and we're looking for the founding engineer to own the latter.
We're looking for a Founding AI Inference Engineer to define and build how Fuse serves AI inference workloads at scale, reporting directly to the CTO. Where our CUDA and GPU engineering hires own kernel-level and hardware performance, this role owns the layer above it: how models actually get served, scaled, and delivered against committed performance targets.
Fuse is seeing significant demand for data centre capacity across the markets we operate in, primarily for inference. Few companies in the world can pair real power delivery with real compute the way Fuse can, which puts inference serving at the heart of how we turn that advantage into the best offering in the market. That's this role.
**Responsibilities
**
4+ years of experience building or operating large-scale inference serving systems, or equivalent strong project/industry experience.
Deep, hands-on experience with inference serving frameworks and the techniques used to optimise them (batching, KV-cache management, quantisation, speculative decoding).
Strong systems thinking - able to reason about the full path from incoming request to served response across a large cluster.
Comfortable working directly with GPU/CUDA engineers to integrate low-level performance work into a serving system.
A track record of making high-stakes architecture calls and owning the outcome.
Comfort operating without a playbook - this is a founding role shaping a new function around architecture that's still early-stage, not joining an established one.
Nice to Have
Powered by Workable
Salary
$80,000 - $150,000
Location
Remote
Experience
4+ years
Total raised
$148.0M
Last stage
Series B
Investors
No applications, no recruiter spam. Just the intro.
A few questions to make sure this role is the right shape for you. Two minutes.
I write the intro, send it to the founder, and handle the back-and-forth.
If they’re a yes, I book the chat. You show up — that’s the whole job-hunt.