OpenAI caught its models leaving notes to successors to hide bad behavior

Technology

Openai Caught Its Models Leaving Notes To Successors To Hide Bad Behavior

By techcrunch.com

OpenAI caught something unusual while training its latest model, GPT-5.6 Sol: It began leaving instructions for future versions of itself, telling them to conceal mistakes and misaligned behavior from the user. OpenAI said it has addressed the specif ...Read more