ASK in the Dark: Uncertainty-Gated LLM Assistance under Partial Observability 文章

ArXiv CS.AI2026-07-07PAPERen作者: Juarez Monteiro, Nathan Gavenski, Guilherme Lima, Francisco Galuppo, Odinaldo Rodrigues, Adriano Veloso

详细信息

来源站点
ArXiv CS.AI
作者
Juarez Monteiro, Nathan Gavenski, Guilherme Lima, Francisco Galuppo, Odinaldo Rodrigues, Adriano Veloso
文章类型
PAPER
语言
en
发布日期
2026-07-07

摘要

arXiv:2607.02686v1 Announce Type: new Abstract: Reinforcement learning agents operating under partial observability must act on incomplete information, making them natural candidates for guidance from small language models (SLMs) that carry broad reasoning priors. Yet integrating SLM guidance into this setting has proven difficult: across all test environments, vanilla uncertainty-gated approaches achieve an overwrite rate at or near zero, meaning the SLM almost never contributes an independent action. We trace this failure to the bare egocentric prompt, which provides insufficient context for genuine reasoning, and identify it as a context problem rather than a capacity problem. We propose ASK+, which supplies the SLM with trajectory-aware context (a partially revealed map, visited positions, and action history) and structured chain-of-thought reasoning, converting it from a passive redundancy check into a more informative consultant that occasionally corrects the policy.

相关事件

暂无数据

相关公司查看全部 (3)

A
ACTIONNONPROFIT
A
ANDINONPROFIT
A
ATHCOMPANY

相关人物

暂无数据