When LLMs Learn to Be Consistently Wrong: A Multi-Model Study of Linear Representations of Synthetic Deception 文章

ArXiv CS.AI2026-06-01NEWSen作者: Vahideh Zolfaghari

When LLMs Learn to Be Consistently Wrong: A Multi-Model Study of Linear Representations of Synthetic Deception · 相关事件

暂无数据