
Latent Space: The AI Engineer Podcast · Latent.Space
Artificial Analysis: Independent LLM Evals as a Service — with George Cameron and Micah-Hill Smith
·1 hr 18 min·7 clips
AI labs sometimes give special endpoints for benchmarking that might not match what users actually get.
As heard by us
Independent AI benchmarking turns into a business story, then into a debate about what model scores really measure.
Artificial Analysis reads like an independent third-party comparison site for model quality and throughput, with model and hosting provider breakdowns that give it a practical, infrastructural feel.
Why you'd press play
You get the business model, the eval math, and the part where everyone argues with the numbers.
Listen to the show on