StakeBench: Evaluating Language Understanding Grounded in Market Commitment 事件
PRODUCT_LAUNCH2026-05-26影响: MEDIUM
StakeBench: Evaluating Language Understanding Grounded in Market Commitment arXiv:2605.26074v1 Announce Type: new Abstract: Existing financial NLP benchmarks often rely on labels supplied by outside observers, measuring how language is perceived rather than what speakers have committed to in the market. We introduce StakeBench, an evaluation framework for language understanding grounded in market commitment. StakeBench links 560,876 comments from 2,261 resolved markets to verified position, act
相关产品查看全部 (10)
相关报道查看全部 (1)
StakeBench: Evaluating Language Understanding Grounded in Market Commitment
ArXiv CS.CL2026-05-26