OmniDelta: Skill-Driven Budget Allocation for Token Compression in OmniLLMs
DGX agentarXiv:2607.25669v1 Announce Type: new Abstract: Emerging Omni-modal Large Language Models (OmniLLMs) enable unified understanding of text, audio, and video, but their long audio-video token sequences