14
you are viewing a single comment's thread
view the rest of the comments
[-] BioMan@awful.systems 2 points 2 days ago* (last edited 2 days ago)

Machine learning training has always presented as a power law, with exponential increases in processing required for linear increases in performance. The Chinchilla scaling laws paper and efficient compute frontier papers let them select the optimal tradeoff between number of parameters and how long to cook it, improving performance greatly by letting you predict how to best use X amount of computation for Y amount of time, setting off people spending hundreds of millions of dollars since they actually knew they could use it optimally, directly leading to the perceived massive increase in performance from 2022-2024ish as they had a one-time burst of knowing how to optimally partition computation and convince people to spend lots of money at once. Everything since that time has been exponentially diminishing returns as expected.

this post was submitted on 14 Sep 2026
14 points (100.0% liked)

TechTakes

2698 readers
214 users here now

Big brain tech dude got yet another clueless take over at HackerNews etc? Here's the place to vent. Orange site, VC foolishness, all welcome.

This is not debate club. Unless it’s amusing debate.

For actually-good tech, you want our NotAwfulTech community

founded 3 years ago
MODERATORS