HugMap
人工智能
云计算
半导体
网络安全
企业软件
区块链
量子计算
生物科技
新能源与智能制造
智能穿戴
机器人
智能手机
图谱探索
趋势分析
登录
注册
How Does Alignment Tuning Shape Representations of Sycophancy and Related Cue-Induced Biases in LLMs?
文章
ArXiv CS.CL
2026-07-21
PAPER
en
作者: Prakhar Gupta, Terry Jingchen Zhang, Florent Draye, Bernhard Sch\"olkopf, Zhijing Jin
查看原文
→
关系图谱
概览
相关事件
相关公司
相关人物
相关产品
相关技术
How Does Alignment Tuning Shape Representations of Sycophancy and Related Cue-Induced Biases in LLMs? · 相关技术
相关技术
pretraining
sycophancy
hidden states
causal intervention
leave-one-dataset-out transfer
Language model probing
alignment tuning
CBCT
cue-induced biases
LLM