Independent AI evaluations lab
We build independent and contamination-proof benchmarks that measure real world performance.
LLM Stats is the most complete LLM leaderboard. We have the most complete archive of LLM benchmark results and also run independent evaluations that are not the classical ones that are already in the training data of most models.
Our mission: become the biggest community dedicated to AI transparency.
Last stage
Seed
Investors
Sebastian Crossa
Co-Founder @ LLM Stats. Previous founding engineer at Micro building the future of email (backed by a16z), as well as founding engineer at Atrato Pago (W21). Formerly built and scaled Minecraft servers during my spare time during highschool.
LinkedInJonathan Chávez
Co-Founder of LLM Stats. Early employee on the LLM observability team at Datadog. Conducted undergrad research on vision transformers for particle physics and RL for robotics.
LinkedInNo applications, no recruiter spam. Just the intro.
A few questions to make sure this role is the right shape for you. Two minutes.
I write the intro, send it to the founder, and handle the back-and-forth.
If they’re a yes, I book the chat. You show up — that’s the whole job-hunt.