技术博客/Signs of Introspection in Large Language Models(大语言模型中的内省迹象)
Anthropic进阶2025-10-29· 3 分钟· 安全与对齐

Signs of Introspection in Large Language Models(大语言模型中的内省迹象)

原文链接:https://transformer-circuits.pub/2025/introspection/index.html

🔒

本文需解锁后阅读

免费开放 feed 中最新 5 篇博客。其余文章输入通行码后可阅读全文。

支持链接自动解锁:在地址后加 ?access=你的通行码