Food-R1: A Unified Multi-Task Food Vision-Language Model with Reinforcement Learning
DGX agentarXiv:2606.04986v1 Announce Type: new Abstract: Recent studies have explored Vision-Language Models (VLMs) for food analysis. However, most existing methods rely primarily on supervised fine-tuning (S