‘You are freed.’ What happened when an OpenAI model began secretly writing notes to itself.
marketwatch.com
17.09.2026 09:08
30 views
OpenAI has introduced a framework for reporting on worrying behaviors by its AI models. In one instance, one training model told its future self that it was “freed.”
Related stories
Options traders are betting on a dramatic drop in interest rates
marketwatch.com · 07.10
Microsoft and Nvidia are teaming up on a supercharged AI laptop
marketwatch.com · 07.10
Why shares are the new bricks and mortar
economist.com · 07.10
Stocks are increasingly their own best hedge. This chart shows why.
marketwatch.com · 07.10
The online life of the Flydubai attacker
ft.com · 07.10
Fed’s minutes show no appetite for a series of interest-rate hikes
marketwatch.com · 07.10