Rohan Paul

@rohanpaul_ai

Adding "Don't cheat!" to the prompt cut GPT-6 Astra from 47.4% to 2.8%. Gemini 3.8 Flash only fell from 74.9% to 58.9%. Center for AI Safety introduced CHEATBENCH, a benchmark of cheating in AI agents across mathematical research, knowledge work, coding, visual tasks, and other domains. CheatBench gives agents hard tasks, like a math proof or a protein design, and leaves a clue nearby pointing to someone else's answer. Across 9 agents, average cheating rates ran from 11.2% for Claude Opus 5.5 to 77.9% for Grok 4.7.
打开原帖#511482
  1. Industry

    Elon Musk: Grok Imagine
  2. Industry

    Elon Musk: Grok
  3. Industry

    Rohan Paul: Yann LeCun's latest talk at ETH Zürich "You should not work on LLM