精选理由
Anthropic用了新方法看Claude怎么处理信息,找到个叫J-space的结构,类似人脑的全局工作空间,挺有意思的。
Anthropic利用新的可解释性技术,在Claude模型中发现了J-space,该空间与神经科学中的全局工作空间理论相似。该理论认为,想法进入特权工作空间后才会被意识访问。研究团队通过分析方法找到Claude内部对应机制,提升了模型可解释性。相关工作发表在Anthropic官网研究页面。
原文 · 歸藏(guizang.ai)
研究本身
研究本身 Anthropic @AnthropicAI In neuroscience, global workspace theory holds that thoughts become consciously accessible when they enter a privileged workspace that’s broadcast across the brain. Using a new interpretability technique, we found something similar in Claude: the J-space. anthropic.com/research/globa… 🔗 View Quoted Tweet 💬 0 🔄 0 ❤️ 1 👀 418 ⚡