Less is More: Compact-Token Masked Feature Prediction for Skeleton Representation Learning
arXiv:2603.10648v3 Announce Type: replace Abstract: Current skeleton representation learning paradigms face distinct limitations: Contrastive Learning (CL) often overlooks fine-grained motion details,