Incentivizing Temporal-Awareness in Egocentric Video Understanding Models
DGX agentMultimodal large language models (MLLMs) have recently shown strong performance in visual understanding, yet they often lack temporal awareness, particularly in egocentric settings where reasoning dep