SMRTR AIAug 18, 2025Hacker News

The lottery ticket hypothesis: why neural networks work

The lottery ticket hypothesis: why neural networks work

SMRTR summary

Breaking centuries of accepted theory, today's most powerful AI systems succeed by defying what was once considered mathematical law. Five years ago, training neural networks with trillions of parameters would have been dismissed as foolish.

"Bigger models just overfit," was the mantra, backed by the bias-variance tradeoff principle that had governed learning systems for over 300 years. The logic seemed unassailable: make your model too simple, it misses patterns; make it too complex, it memorizes noise instead of signals.

Then in 2019, researchers committed scientific heresy - they scaled neural networks far beyond the point where theory predicted catastrophic failure. Instead of collapsing, these massive models showed "double descent," where after initially overfitting, performance dramatically improved again.

The explanation came from MIT's "lottery ticket hypothesis": large networks succeed not by memorizing, but by containing countless potential simple solutions with different starting conditions. Training becomes a massive lottery draw where the best-initialized small network emerges victorious.

This revelation reconciles empirical success with classical theory. Intelligence isn't about memorizing complexity - it's about finding elegant patterns that explain complex phenomena. Scale simply provides more lottery tickets, more chances to find those optimal solutions.

SMRTR provides this summary for quick context. The original article belongs to Hacker News.

Read the original article
SMRTR AI

Get the next batch of curated stories in your inbox.

This archive is built from SMRTR newsletter stories. Subscribe for hand-picked stories without the extra noise.

Related Stories

Browse AI
AIAug 27, 2026

SpaceKit AI

SpaceKit AI is a public research hub for Growformer, an ML system with a promote-freeze neural substrate, language model experiments, neural cellular automata, and reproducible...

AIAug 27, 2026

Small Models Have Arrived

Small AI models are now 10x cheaper than before, making high-volume consumer and business AI apps financially viable—tasks costing $1 now cost ~$0.10, enabling affordable,...

AIAug 27, 2026

Why I Am Right About AI

The piece satirizes tech-bro certainty about AI replacing writers, using absurd logic and fake statistics to parody real thought leaders—exposing how hollow and self-serving those...