🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production
-
Updated
Aug 16, 2024 - Python
🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production
Unofficial Parallel WaveGAN (+ MelGAN & Multi-band MelGAN & HiFi-GAN & StyleMelGAN) with Pytorch
Persian/Farsi text to speech(TTS) training using coqui tts
Include Basis-MelGAN, MelGAN, HifiGAN and Multiband-HifiGAN, maybe NHV in the future.
🎙️ Arabic TTS models (Tacotron2, FastPitch)
Speech synthesis (TTS) in low-resource languages by training from scratch with Fastpitch and fine-tuning with HifiGan
zero-shot realtime TTS system, fully offline, free and open source
Ultrafast GAN based Vocoder for Text to Speech
🎙️ Arabic TTS models (FastPitch, Mixer-TTS) in the ONNX format — Python package for offline speech synthesis 🚀📦
HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis
SA-toolkit: Speaker speech anonymization toolkit in python
Fast and Small codec for flow-matching models
RADTTS + HiFiGAN vocoder
Zero shot singing voice conversion, whats better to make than this on a Monday?
🇺🇦 Ukrainian RAD-TTS++ models (decoder + models with 3 voices) and HiFiGAN model
To associate your repository with the hifigan topic, visit your repo's landing page and select "manage topics."