Perplexity

@perplexity_ai

New research: We post-trained a Computer model to learn from its own errors using hint-guided self-distillation. In a live A/B test, a later trained checkpoint reduced tool-call failures by 21.2% relative to an earlier checkpoint. https://t.co/3MFrp1yxDt
打开原帖#511482
  1. Industry

    Alexandr Wang: we are considering opening a muse merch store i am doing market resea…
  2. Industry

    Alexandr Wang: another one!
  3. Industry

    Elon Musk: Interesting