Rohan Paul
@rohanpaul_ai
Yann LeCun's (@ylecun ) latest talk at ETH Zürich
Scaling LLMs to reach AGI is "impossible"
A large language model is trained on about 30 trillion tokens, which is roughly 10^14 bytes of text and would take a person about 400,000 years to read.
A 4-year-old child receives about the same amount of data, 10^14 bytes, through vision alone in about 1 year and 10 months.
In his view, intelligence is the ability to learn new tasks quickly or perform them without prior training, as a teenager learns to drive in about 20 hours. Scaling increases stored knowledge, but it does not produce this ability to adapt.
----
From "Perfology Clips" YouTube channel, (link in comment)