Skip to content
View zge's full-sized avatar

Block or report zge

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

PyTorch re-implementation of Speech-Transformer

Python 102 33 Updated Nov 19, 2021

Robust Speech Recognition via Large-Scale Weak Supervision

Python 107,347 13,043 Updated Jul 28, 2026

A resource for learning about Machine learning & Deep Learning

Python 8,484 2,787 Updated Aug 17, 2024

Trax — Deep Learning with Clear Code and Speed

Python 8,309 820 Updated Sep 26, 2025

Tutorials on implementing a few sequence-to-sequence (seq2seq) models with PyTorch and TorchText.

Jupyter Notebook 5,707 1,355 Updated Jan 20, 2024

A scalable generative AI framework built for researchers and developers working on Large Language Models, Multimodal, and Speech AI (Automatic Speech Recognition and Text-to-Speech)

Python 18,137 3,557 Updated Aug 16, 2026

CommonMark spec, with reference implementations in C and JavaScript

Python 5,133 353 Updated Apr 27, 2026

Source code complementing our paper for acoustic event classification using convolutional neural networks.

Python 71 28 Updated Jan 31, 2021

Grapheme to phoneme conversion with deep learning.

Python 435 63 Updated Dec 8, 2023

Audio super resolution using neural networks

Python 1,263 211 Updated Oct 24, 2023

The official implementation of the Interspeech 2021 paper WSRGlow: A Glow-based Waveform Generative Model for Audio Super-Resolution.

Python 127 19 Updated Sep 7, 2021

End-to-End Neural Diarization

Python 436 63 Updated Aug 30, 2021

A curated list of awesome Speaker Diarization papers, libraries, datasets, and other resources.

1,892 241 Updated Aug 12, 2026

Implementation of "MOSNet: Deep Learning based Objective Assessment for Voice Conversion"

Python 379 64 Updated Jul 21, 2024

End-to-End Speech Processing Toolkit

Python 9,924 2,421 Updated Aug 14, 2026

Large, modern dataset for speech recognition

Shell 732 67 Updated Feb 26, 2024

Self-Supervised Speech Pre-training and Representation Learning Toolkit

Python 2,561 534 Updated Mar 12, 2026

A series of convenience functions to make basic image processing operations such as translation, rotation, resizing, skeletonization, and displaying Matplotlib images easier with OpenCV and Python.

Python 4,592 1,023 Updated Jun 24, 2024

Trainable algorithm for accurate force alignment

Rust 5 Updated Oct 19, 2015