The Lipreading Gap: Do VSR Models Perceive Visual Speech Like Human Lipreaders?
arXiv:2606.07435v1 Announce Type: cross Abstract: Visual speech recognition (VSR) models now surpass human lipreaders on benchmarks, but do such gains establish human-like visual speech perception? To