13 September 2026
Heard In AI

Tag

AI benchmarks

Articles about AI benchmarks from podcasts, articles and papers, with links to the original sources.

Astra’s robot-arm gains stop short of precision

RoboCurve reports that GPT-6 Astra placed a block in a bowl in 19 of 20 trials, up from 8 of 20 for its predecessor. But it completed only two puzzle insertions—the same as the older model—in a 120-trial evaluation that separates basic manipulation gains from reliable precision.

3 min read

Better data beat model recipes in a small-scale training test

A controlled comparison of 2019–2025 datasets and training recipes reported compute-efficiency gains of 12-fold from data improvements versus 3.7-fold from model recipes. The Moonshots panel explored the business opportunity—and used BloombergGPT’s reportedly short-lived advantage to question whether owning unique data is enough.

5 min read