-
Wuhan University
- Wuhan, China
- https://orcid.org/0000-0002-5339-7182
Stars
Official implementation of Lombard-VLD (IEEE S&P 25)
Official Implementation of VoxTracer (MM' 23)
⏰ Agenticly track worldwide conference deadlines (Website, Python Cli, Wechat Applet)
Demonstrate all the questions on LeetCode in the form of animation.(用动画的形式呈现解LeetCode题目的思路,完整单步/回看/变速/语音讲解在 algomooc.com)
Self-Supervised Speech Pre-training and Representation Learning Toolkit
Deep neural networks for voice conversion (voice style transfer) in Tensorflow
"Deep Generative Modeling": Introductory Examples
A Non-Autoregressive Text-to-Speech (NAR-TTS) framework, including official PyTorch implementation of PortaSpeech (NeurIPS 2021) and DiffSpeech (AAAI 2022)
VITS: Conditional Variational Autoencoder with Adversarial Learning for End-to-End Text-to-Speech
An implementation of Microsoft's "FastSpeech 2: Fast and High-Quality End-to-End Text to Speech"
Source code for paper "Who is real Bob? Adversarial Attacks on Speaker Recognition Systems" (IEEE S&P 2021)
kaldi-asr/kaldi is the official location of the Kaldi project.
This is a pytorch implementation of the paper: StarGAN-VC: Non-parallel many-to-many voice conversion with star generative adversarial networks
StarGANv2-VC: A Diverse, Unsupervised, Non-parallel Framework for Natural-Sounding Voice Conversion
Global Rhythm Style Transfer Without Text Transcriptions
Official implementation of the SPL paper "One-class Learning Towards Synthetic Voice Spoofing Detection"