François Chollet

@fchollet

What if the jagged frontier is mainly math + code (which you can push arbitrarily far with RLVR), and everything else starts to plateau because it is still bottlenecked by human generated data? Model performance in non-verifiable areas has kept improving steadily, albeit much slower than for math and code. But is that steady improvement a side effect of a higher G (itself driven by RLVR), or only a function of the amount of new human data getting injected into training (which is still continually happening on a massive scale)? A lot of things depend on the answer to this question
打开原帖#511482
  1. Industry

    Greg Brockman: towards acceleration of scientific discovery and improving quality of…
  2. Research

    François Chollet: Does G "emerge" from math + code RLVR?
  3. Industry

    Replit ⠕: Join us tomorrow at #SFTechWeek with @passionfrootme @elevenlabs and…