Safety
easyaligner: Forced alignment with GPU acceleration and flexible text normalization (compatible with all w2v2 models on HF Hub) [P]
easyaligner is a forced alignment library designed to be performant and easy to use , leveraging GPU acceleration to align audio with text transcriptions. The tool supports flexible text normalization
easyaligner is a forced alignment library designed to be performant and easy to use , leveraging GPU acceleration to align audio with text transcriptions. The tool supports flexible text normalization and is compatible with all Wav2Vec2 models available on the Hugging Face Hub, enabling efficient phoneme-level alignment for speech processing tasks.
Related
- Gradient boundaries through confidence intervals for forced alignment estimates using model ensembles
- MM-LIMA: Less Is More for Alignment in Multi-Modal Datasets
- Descriptor-Injected Cross-Modal Learning: A Systematic Exploration of Audio-MIDI Alignment via Spectral and Melodic Features
Source: r/MachineLearning | 2026-04-18