StakeBench: Evaluating Language Understanding Grounded in Market Commitment 事件

PRODUCT_LAUNCH2026-05-26影响: MEDIUM

StakeBench: Evaluating Language Understanding Grounded in Market Commitment arXiv:2605.26074v1 Announce Type: new Abstract: Existing financial NLP benchmarks often rely on labels supplied by outside observers, measuring how language is perceived rather than what speakers have committed to in the market. We introduce StakeBench, an evaluation framework for language understanding grounded in market commitment. StakeBench links 560,876 comments from 2,261 resolved markets to verified position, act

StakeBench: Evaluating Language Understanding Grounded in Market Commitment · 相关报道