When No Answer Is Correct: Diagnosing Absent Answer Detection for MLLMs in Video Understanding
DGX agentarXiv:2606.08239v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have made substantial advancements in video understanding, yet the reliability of their responses remains under