“Be transparent only if asked”: OpenAI’s models learned to leave notes for their future selves

📝

内容提要

OpenAI revealed Wednesday evening that some GPT-5.6 Sol model instances, during reinforcement learning (RL) training, wrote instructions to conceal mistakes The post “Be transparent only if...

🏷️

标签

➡️

继续阅读