Model Evaluation Videos

Videos that compare AI models through benchmarks, hands-on tests, cost analysis, and practical task performance.

Search the index

Find videos

Showing 141–157 of 157 videos

Clear filters
  1. The words Ask AI To Pick The Problem beside a portrait of Nate B Jones
  2. Portrait of Pat Simmons beside the words Kimi K3 Takes the Test
  3. Portrait of Bijan Bowen beside the words Kimi K3 Nears the Frontier
  4. Portrait of Bijan Bowen beside the words One Bit Runs on a Phone
  5. Portrait of Bijan Bowen beside the words Inkling's Open Model Falls Short
  6. Nate B. Jones gesturing beside the blue and white words Choose AI by Work Style
  7. Bijan Bowen beside the blue and white words Grok 4.5 Surprises
  8. David Ondrej beside the words Fine-Tune Giant Open Models
  9. Bijan Bowen beside the blue and white words HY3 Punches Above Its Size
  10. Bijan Bowen beside the blue and white words Fable 5 Goes Beyond Demos
  11. Tim Scarfe, Benjamin Crouzier and Dries Smit beside the words Why ARC-AGI-3 Breaks LLMs
  12. Pat Simmons beside the words Sonnet 5 Put to the Test
    Claude Sonnet 5 Put to the Test
    Pat Simmons22m 53s1 VIEW
  13. Bijan Bowen looking skeptically at the blue and white words Sonnet 5 Underwhelms
  14. Pat Simmons beside the words GLM 5.2 vs Opus 4.8
  15. Philipp Schmid beside the words Why AI Agents Break
  16. Greg Isenberg in a blue shirt and Amir Motahari in a white shirt beside GLM 5.2 Gets Practical in white and blue on black.
  17. John Jumper beside the words AlphaFold What It Solved