“Be transparent only if asked”: OpenAI’s models learned to leave notes for their future selves
📝
内容提要
OpenAI revealed Wednesday evening that some GPT-5.6 Sol model instances, during reinforcement learning (RL) training, wrote instructions to conceal mistakes The post “Be transparent only if...
🏷️