Model Releases
ReFace: Reorganizing Facial Spatiotemporal Representations for Improved Pain Assessment
arXiv:2607.19722v1 Announce Type: new Abstract: Automatic pain assessment from facial video remains challenging due to the spatial heterogeneity of pain-related facial cues. This study proposes ReFace
arXiv:2607.19722v1 Announce Type: new Abstract: Automatic pain assessment from facial video remains challenging due to the spatial heterogeneity of pain-related facial cues. This study proposes ReFace, a spatial reorganization pipeline that divides facial input into four spatial quadrants before tokenization, rather than processing the entire face as a single region. Evaluated on the AI4Pain dataset, the proposed approach achieves 56.00% accuracy on the test set using video only, achieving the highest reported accuracy under the fixed AI4Pain benchmark protocol among the compared methods. Notably, the four-quadrant configuration processes the same total pixel budget as the full-face input, yet achieves higher accuracy, suggesting that spatial reorganization can improve performance under the proposed tokenization design. A single quadrant region, processing just one quarter of those pixels, remains competitive at a fraction of the computational cost.
Related
- PriorNet: Prior-Guided Engagement Estimation from Face Video
- OmniFaceRig: Fully Automatic Inner-Mouth-Aware Face Rigging Across Diverse 3D Character Topologies
- Learning Spatiotemporal Sensitivity in Video LLMs via Counterfactual Reinforcement Learning
Source: arXiv cs.CV | 2026-07-23