Pinned Loading
Repositories
- ConsistentTTS Public Forked from latentforge/VoiceStudio
[ICASSP 2027] ConsistentTTS - Speaker Stabilization Method for Voice Design TTS
- LibriSONA Public Forked from latentforge/VoiceStudio
[ICASSP 2027] LibriSONA Benchmark for TTS Models (Metric & Dataset)
- transformers-tts Public Forked from huggingface/transformers
🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.
- transformers-moshi Public Forked from huggingface/transformers
🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.
- vocos Public Forked from gemelo-ai/vocos
Vocos: Closing the gap between time-domain and Fourier-based neural vocoders for high-quality audio synthesis
- EmergentTTS-Eval-public Public Forked from boson-ai/EmergentTTS-Eval-public
[NeurIPS' 25] Benchmark for evaluating TTS models on complex prosodic, expressiveness, and linguistic challenges.
- CausVid Public Forked from tianweiy/CausVid
(CVPR 2025) From Slow Bidirectional to Fast Autoregressive Video Diffusion Models
Top languages
Loading…
Most used topics
Loading…