SVHalluc: Benchmarking Speech-Vision Hallucination in Audio-Visual Large Language Models 事件

Name: SVHalluc: Benchmarking Speech-Vision Hallucination in Audio-Visual Large Language Models
Start: 2026-06-03

PRODUCT_LAUNCH2026-06-03影响: MEDIUM

SVHalluc: Benchmarking Speech-Vision Hallucination in Audio-Visual Large Language Models arXiv:2606.02642v1 Announce Type: cross Abstract: Despite the success of audio-visual large-language models (LLMs), they can produce plausible but ungrounded outputs, termed hallucination. Existing benchmarks focus on environmental sounds (e.g., dog barking) to indicate event occurrence. In contrast, human speech carries fundamentally different, rich semantics and temporal structures, yet it remains unexplo

人工智能

关系图谱

SVHalluc: Benchmarking Speech-Vision Hallucination in Audio-Visual Large Language Models 事件

相关公司查看全部 (7)

相关人物查看全部 (1)

相关产品查看全部 (10)

相关技术查看全部 (10)

相关报道查看全部 (1)