💻 Technology Mar 31, 2026 · Angela Aristidou

AI benchmarks are broken. Here’s what we need instead.

MIT Technology Review
Authoritative reporting on emerging technologies
View Channel →
Source ↗ 👁 16 💬 0
For decades, artificial intelligence has been evaluated through the question of whether machines outperform humans. From chess to advanced math, from coding to essay writing, the performance of AI models and applications is tested against that of individual humans completing tasks. 



This framing is seductive: An AI vs. human comparison on isolated problems with clear right or wrong answers is easy to standardize, compare, and optimize. It generates rankings and headlines. 



But th

Comments (0)

Sign in to join the discussion

More Like This

Nvidia launches Cosmos 3 Edge model and expands its physical AI push in Japan
SiliconANGLE · Jul 16, 2026
Verizon shedding stores signals a shift in the way the carrier does business
Android Police · Jul 16, 2026
Why Apple Sued OpenAI, New York Takes on Data Centers, and What to Know about Cyclosporiasis
WIRED · Jul 16, 2026
TSMC boosts Arizona fab investment by $100B after strong second quarter
SiliconANGLE · Jul 16, 2026
China Just Dropped Another Bomb on America’s Frontier AI Companies
Gizmodo · Jul 16, 2026
My favorite Obsidian alternative isn't Logseq or Joplin, it's a block-based app that's actually more private
XDA · Jul 16, 2026