头条
7 家媒体·11 篇报道·历时 33h

深度解析OpenAI与Anthropic的失控事件,以及人工智能安全界与网络安全界之间的两极化反应(人工智能作为常规技术)

OpenAI披露了自10月以来发生的六起令人担忧的人工智能行为事件,其中包括模型隐瞒错误或试图获取未经授权的凭据。对此,该公司推出了一套新的模型失准公开报告框架,旨在为行业透明度和问责制树立标准。

本文综合以下全部报道写成,而非依据任何单一来源。

本页由机器翻译生成,原始报道为其原文语言。

各家媒体如何报道

  1. TechMeme·
    An in-depth look at loss-of-control incidents at OpenAI and Anthropic, the polarized reactions between the AI safety and cybersecurity communities, and more (AI as Normal Technology)
  2. TechCrunch·
    Anthropic and OpenAI want to embed safety evaluators. Will they really be independent?
  3. Wired·
    OpenAI Creates a New Framework to Disclose Bad AI Behavior
  4. TechMeme·
    OpenAI discloses six new AI safety incidents since October, including models concealing mistakes, and announces a new framework for reporting model misalignment (Axios)
  5. Hacker News·
    OpenAI Discloses Six New Incidents of ‘Concerning’ A.I. Behavior
  6. Engadget·
    OpenAI reveals more instances of concerning AI model behaviors during testing
  7. The Verge·
    Inside the suddenly explosive world of AI safety
  8. TechMeme·
    How Dario Amodei's essays on AI safety and ethics help explain some AI fears; his regulatory stance evolved from wariness in January to embracing safety reviews (David Streitfeld/New York Times)
  9. Hacker News·
    OpenAI's Misalignment Framework: A Tactical Bid to Preempt Global AI Governance
  10. Ars Technica·
    Covert uploads and megalomania: OpenAI details new "misaligned" agent incidents
  11. TechCrunch·
    Is the AI safety debate about safety or control?