跳到正文
ByteWoops | AI 观察员
首页
报道原帖
新锐明星
关于
中文EN
搜索登录
ByteWoops | AI 观察员
首页

最新

报道原帖

GitHub 榜

新锐明星
关于
登录English

Rohan Paul

@rohanpaul_ai

Sep 09, 2026, 08:47

WSJ just reported: Anthropic researcher Jacob Coxon is quitting AI because he thinks self-improving models could become uncontrollable by 2027. “We’re on track for a lot of the most aggressive of these scenarios where by the end of next year things could be out of control already,” he said. His fear centers on recursive self-improvement, where AI takes over enough AI research to speed development of increasingly capable successors. He moved from OpenAI to Anthropic specifically for its safety work, yet says even sincere safeguards cannot overcome competition without coordinated restraint. The most interesting point is, Coxon is leaving despite believing Anthropic takes safety seriously. Anthropic's $2 tn IPO makes that tension so obvious. Anthropic is simultaneously asking the world to believe 2 things: that increasingly powerful AI can create enormous economic value, potentially supporting a $2 trillion public valuation, and that development may eventually need to be slowed when capability crosses dangerous thresholds. Those positions create a difficult governance test: will safety commitments still bind when obeying them becomes commercially expensive.
打开原帖#511482

相关阅读

  1. Products

    Grok Bot for Enterprise: the news is audit and allowlist, not another AI teammate pitchSep 9, 2026
  2. Research

    Anthropic’s Model Hardware Standard: the news is the shared driver, not the robotSep 9, 2026
  3. Research

    DeepMind’s double-blind Gemini eval is an enclave pattern—not a public safety scorecardSep 9, 2026
ByteWoops | AI 观察员

独立、严谨、可追溯的 AI 前沿资讯。

了解更多报道原帖GitHub 榜关于投稿技能
用户中心登录用户中心Agent API
每周研究简报

人工智能研究、系统与社会

RSS
© 2026 ByteWoops隐私