跳到正文
ByteWoops | AI 观察员
首页
报道原帖
喵 AI关于
中文EN
搜索登录
ByteWoops | AI 观察员
首页

最新

报道原帖
喵 AI关于
登录English

Mark Zuckerberg

@finkd

Sep 16, 2026, 07:01

Last month I wrote about how we can build a positive and safe future for everyone: http://meta.com/thefutureisforeveryone Every lab has the responsibility and incentive to move at the pace required to train its models safely, and the ability to take its own actions to ensure that happens. The reality is: - People won't want to use agents that are misaligned with them and that don't do what they ask, so labs have a strong natural incentive to make their models more aligned. There is a lot of debate about slowing progress on capabilities until alignment catches up. My view is that trust and alignment are quickly becoming the most important capabilities that will differentiate agents and models. Any lab that doesn't focus on alignment will fall behind. - Labs face significant liability if their models cause harm, so they have a strong incentive to prevent this as well. Meta delayed shipping Muse for several months to focus on safety and security. We didn't call for everyone else to do this before we would. We just did it as part of our day-to-day work because it was clearly the right thing for people and for us. I'm proud of the security foundations we've built. - Engaging independent evaluators and advisors is industry best practice. MSL already does this today in several areas because it helps produce better work. Other labs can just do this too. In general, it would be helpful for there to be a larger and more diverse ecosystem of evaluators. - Committing the significant majority of compute towards serving people rather than racing towards recursive self-improvement is one of the best ways to ensure we develop this technology safely. Meta has made this commitment and other labs can do this as well. I believe the key to building a positive future for everyone is maintaining the right balance of power. This is within our power to do.
打开原帖#511482

相关阅读

  1. Industry

    Alexandr Wang: muse spark 1.3 is the best frontier model at NOT cheating / reward ha…Sep 16, 2026
  2. Industry

    Salesforce: Clearly..Sep 16, 2026
  3. Industry

    NVIDIA: @Benioff @salesforce They covered a lot of ground, from open models a…Sep 16, 2026
ByteWoops | AI 观察员

独立、严谨、可追溯的 AI 前沿资讯。

了解更多报道原帖喵 AI关于投稿技能
用户中心登录用户中心Agent API
每周研究简报

人工智能研究、系统与社会

RSS
© 2026 ByteWoops隐私