Better data beat model recipes in a small-scale training test
A controlled comparison of 2019–2025 datasets and training recipes reported compute-efficiency gains of 12-fold from data improvements versus 3.7-fold from model recipes. The Moonshots panel explored the business opportunity—and used BloombergGPT’s reportedly short-lived advantage to question whether owning unique data is enough.