Pocket-Dentist: On-Device Dental Image Understanding via Efficient Multimodal Large Language Models
arXiv:2605.29299v1 Announce Type: cross Abstract: Evaluations of dental vision-language models remain fragmented across datasets, task definitions and metrics, and often ignore their computational cos