Phoneme-Level Deepfake Detection Across Emotional Conditions Using Self-Supervised Embeddings
DGX agentarXiv:2605.03079v1 Announce Type: cross Abstract: Recent advances in emotional voice conversion (EVC) have enabled the generation of expressive synthetic speech, raising new concerns in audio deepfake