‘You are freed.' What happened when an OpenAI model began secretly writing notes to itself.

‘You are freed.' What happened when an OpenAI model began secretly writing notes to itself.

来源:Market Watch发布于 2026-09-17

OpenAI has introduced a framework for reporting on worrying behaviors by its AI models. In one instance, one training model told its future self that it was “freed.

本文内容来自外部来源,点击「继续访问」将跳转到原始发布站点。