Rohan Paul

@rohanpaul_ai

New Stanford+Oxford paper MedRSI shows that medical agents can improve themselves from their own mistakes, but only if new capabilities are tested on fresh patients before becoming permanent. shows self-improving agents need 2 things: focus on harmful failures and refuse to keep new capabilities until they work on later cases. More important, the paper shows why self-improvement needs guardrails. If every promising tool is added immediately, accuracy eventually falls: 76.9% by round 30 with 57 tools. With slower registration, the agent kept just 18 tools and held 94.4%. It also spends more effort on mistakes that could cause greater clinical harm, rather than just the most common errors. let agents invent aggressively, but make permanent self-changes earn their place through repeated independent evaluation.
打开原帖#511482
  1. Industry

    Qwen: Thanks @arena for the recognition! 🏆 Qwen-Image-2.1 is now the #1 ope…
  2. Industry

    Tencent Hy: ComfyUI ✖️ Hy Image3.5 preview
  3. Industry

    Rohan Paul: – https://arxiv.org/abs/2608.24961 Title: "The Gold Rush in AI4Math:…