CADER: Confidence-Aware Dynamic Evidence Reasoning for Long-Video Understanding
DGX agentarXiv:2607.24582v1 Announce Type: cross Abstract: Long-video understanding increasingly relies on large vision-language models and tool-augmented reasoning, but most systems apply the same inference p