about
Hype, Sustainability, and the Price of the Bigger-Is-Better Paradigm in AI (arxiv.org)
6 points by lamename on Sep 25, 2024 | hide | past | pdf | discuss on HN

In plain words: This analysis checks the belief that ever-bigger AI models are always better by tracing how scale, performance and impact relate. It finds gains don't mainly come from size, while computing needs grow faster than performance, driving costs, energy use, and power held by few.

Abstract · Hype, Sustainability, and the Price of the Bigger-is-Better Paradigm in AI

With the growing attention and investment in recent AI approaches such as large language models, the narrative that the larger the AI system the more valuable, powerful and interesting it is is increasingly seen as common sense. But what is this assumption based on, and how are we measuring value, power, and performance? And what are the collateral consequences of this race to ever-increasing scale? Here, we scrutinize the current scaling trends and trade-offs across multiple axes and refute two common assumptions underlying the 'bigger-is-better' AI paradigm: 1) that performance improvements are driven by increased scale, and 2) that all interesting problems addressed by AI require large-scale models. Rather, we argue that this approach is not only fragile scientifically, but comes with undesirable consequences. First, it is not sustainable, as, despite efficiency improvements, its compute demands increase faster than model performance, leading to unreasonable economic requirements and a disproportionate environmental footprint. Second, it implies focusing on certain problems at the expense of others, leaving aside important applications, e.g. health, education, or the climate. Finally, it exacerbates a concentration of power, which centralizes decision-making in the hands of a few actors while threatening to disempower others in the context of shaping both AI research and its applications throughout society.

Gaël Varoquaux, Alexandra Sasha Luccioni, Meredith Whittaker
arXiv:2409.14160 · cs.CY · submitted Sep 21, 2024 · updated Mar 1, 2025
abstract · pdf · html

add comment on HN
Also discussed: Jun 2025 (3 points, 0 comments)