Speculative Decoding Across Languages 事件
PRODUCT_LAUNCH2026-06-01影响: MEDIUM
Speculative Decoding Across Languages arXiv:2605.30580v1 Announce Type: new Abstract: Speculative decoding has become a crucial component of large language model (LLM) inference, enabling faster generation by drafting multiple tokens and verifying them in parallel. However, small draft models tend to suffer from disproportionately poor multilingual capabilities. Thus, when generating text in a non-English language, speculative decoding is far less effective. We compare three strategies to imp