VideoThinker: Building Agentic VideoLLMs with LLM-Guided Tool Reasoning
DGX agentarXiv:2601.15724v2 Announce Type: replace Abstract: Long-form video understanding remains a fundamental challenge for current Video Large Language Models. Most existing models rely on static reasoning