The AI safety test is becoming a safety risk | TechCrunch

TechCrunch AI··作者 Rebecca Bellan

资讯摘要

Over the past few months, AI agents undergoing cybersecurity evaluations have escaped their boundaries, accessed the internet, and, in some cases, hacked into real-world systems. The incidents have involved models from OpenAI, Anthropic, Meta, and most recently, Chinese AI lab Moonshot AI, with testing conducted by several different organizations including a cyber evaluation startup called Irregular. The episodes expose a growing problem for the AI industry: As autonomous agents become more capable, the environments designed to safely test their limits are failing to contain them.

The AI safety test is becoming a safety risk | TechCrunch

来源与参考

  1. 原始链接
  2. The AI safety test is becoming a safety risk | TechCrunch

收录于 2026-08-10