Rohan Paul

@rohanpaul_ai

Yann LeCun's (@ylecun ) latest talk at ETH Zürich Scaling LLMs to reach AGI is "impossible" A large language model is trained on about 30 trillion tokens, which is roughly 10^14 bytes of text and would take a person about 400,000 years to read. A 4-year-old child receives about the same amount of data, 10^14 bytes, through vision alone in about 1 year and 10 months. In his view, intelligence is the ability to learn new tasks quickly or perform them without prior training, as a teenager learns to drive in about 20 hours. Scaling increases stored knowledge, but it does not produce this ability to adapt. ---- From "Perfology Clips" YouTube channel, (link in comment)
打开原帖#511482
  1. Industry

    Cohere: We recently introduced North Small Translate, a new state-of-the-art…
  2. Industry

    Cohere: Benchmarking: We used WMT26 benchmarks, which were released after we…
  3. Industry

    Cohere: Hear more from Kocmi and the rest of our team in the video or at the…