CG-MLLM: Captioning and Generating 3D content via Multi-modal Large Language Models
DGX agentarXiv:2601.21798v2 Announce Type: replace Abstract: Large Language Models(LLMs) have revolutionized text generation and multimodal perception,but their capabilities in 3D content generation remain und