Stars
🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.
Manipulate audio with a simple and easy high level interface
Any-to-any voice conversion by end-to-end extracting and fusing fine-grained voice fragments with attention
An implementation of Microsoft's "FastSpeech 2: Fast and High-Quality End-to-End Text to Speech"
UHop: An Unrestricted-Hop Relation Extraction Framework for Knowledge-Based Question Answering
Implementation of the paper Tree Transformer
Pytorch Implemetation for our NAACL2019 Paper "Riemannian Normalizing Flow on Variational Wasserstein Autoencoder for Text Modeling" https://arxiv.org/abs/1904.02399
Salute to a Facebook Fanpage Tsai Bro Pro Betel Nut (臉書:財哥專業檳榔攤)
PyTorch implementation of Tacotron speech synthesis model.
📺 Stream YouTube videos as ascii art in the terminal!
A PyTorch implementation of the WaveGlow: A Flow-based Generative Network for Speech Synthesis
This is an open source project (formerly named Listen, Attend and Spell - PyTorch Implementation) for end-to-end ASR implemented with Pytorch, the well known deep learning toolkit.
PyTorch implementation of AVF
Wheel of tensorflow build for cuda9.2 and python3.6
Read, write, and manipulate Praat TextGrid files with Python
Scripts with example usage of tensorflow profiler
A tensorflow implementation of the "Style Tokens: Unsupervised Style Modeling, Control and Transfer in End-to-End Speech Synthesis"
This repository contains the code to reproduce the core results from the paper "Scalable Factorized Hierarchical Variational Autoencoders"
A TensorFlow Implementation of Tacotron: A Fully End-to-End Text-To-Speech Synthesis Model
Deep neural networks for voice conversion (voice style transfer) in Tensorflow
Python Audio Analysis Library: Feature Extraction, Classification, Segmentation and Applications
Speech Enhancement Generative Adversarial Network in TensorFlow