
Latent Space: The AI Engineer Podcast · Latent.Space
[State of Code Evals] After SWE-bench, Code Clash & SOTA Coding Benchmarks recap — John Yang
·18 min·1 clip
In CodeClash, AI models maintain their own code bases, edit them each round, and then face off in an arena to see whose code performs better.
Listen to the show on