Skip to content
View cst781's full-sized avatar

Block or report cst781

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Self-Supervised Speech Pre-training and Representation Learning Toolkit

Python 2,561 534 Updated Mar 12, 2026

The GitHub repository for the paper "Informer" accepted by AAAI 2021.

Python 6,514 1,298 Updated Jun 20, 2025

Python implementation of performance metrics in Loizou's Speech Enhancement book

Python 458 93 Updated Feb 15, 2025

A collection of filters for real-time audio processing

Rust 54 3 Updated May 25, 2021

Real-time voice-changer for voice-chat, etc. Will support many different voice-filters and features in the future. 🎵

Python 1,041 85 Updated Jul 2, 2023

A first-of-its-kind acoustic simulation platform for audio-visual embodied AI research. It supports training and evaluating multiple tasks and applications.

Python 470 74 Updated Sep 29, 2023

Muzic: Music Understanding and Generation with Artificial Intelligence

Python 4,947 499 Updated Aug 5, 2026

Code for the AAAI 2022 paper "SSAST: Self-Supervised Audio Spectrogram Transformer".

Python 430 65 Updated Aug 14, 2022

Code for the paper "MULTI-BAND MASKING FOR WAVEFORM-BASED SINGING VOICE SEPARATION" that was accepted on EUSIPCO2022

Python 15 4 Updated Jun 18, 2022

Crack LeetCode, not only how, but also why.

Markdown 135,358 23,573 Updated Feb 28, 2026

The implementation of "Dual-branch Attention-In-Attention Transformer for single-channel speech enhancement"

Python 126 21 Updated Jun 29, 2022

A PyTorch implementation of Conv-TasNet described in "TasNet: Surpassing Ideal Time-Frequency Masking for Speech Separation" with Permutation Invariant Training (PIT).

Python 769 156 Updated Apr 6, 2023

Pytorch: Channel-wise subband (CWS) input for better voice and accompaniment separation

Python 102 15 Updated Nov 12, 2021

Pytorch implementation of subband decomposition

HTML 93 13 Updated Jul 26, 2022

A PyTorch implementation of Time-domain Audio Separation Network (TasNet) with Permutation Invariant Training (PIT) for speech separation.

Python 125 31 Updated Jan 27, 2019

🎵 a new stem dataset for Music Demixing research, from the OnAir royalty-free music project

37 4 Updated Mar 14, 2023

Online Normalization for Training Neural Networks (Companion Repository)

Python 89 21 Updated Apr 22, 2021

Perceptual Quality Estimator for speech and audio

C++ 919 143 Updated May 17, 2025

🚀Clone a voice in 5 seconds to generate arbitrary speech in real-time

Python 36,910 5,191 Updated Mar 3, 2026

Music Source Separation; Train & Eval & Inference piplines and pretrained models we used for 2021 ISMIR MDX Challenge.

Python 119 20 Updated Oct 16, 2022
Python 51 8 Updated May 16, 2021

Implementation of "A Deep Learning Loss Function based on Auditory Power Compression for Speech Enhancement" by pytorch

Python 28 5 Updated Jan 31, 2022

Open-Unmix - Music Source Separation for PyTorch

Python 1,501 205 Updated Jun 17, 2024

Deep learning based speech source separation using Pytorch

Jupyter Notebook 319 47 Updated Nov 20, 2020

This repo contains the scripts, models, and required files for the Deep Noise Suppression (DNS) Challenge.

Python 1,459 455 Updated Jul 25, 2024

Multi-Task Audio Source Separation, Two-Stage Model, Complex Domain.

Python 97 13 Updated Jun 13, 2023

The PyTorch-based audio source separation toolkit for researchers

Python 2,583 448 Updated May 13, 2026

implementation of "DCCRN-Deep Complex Convolution Recurrent Network for Phase-Aware Speech Enhancement" by pytorch

Python 1 Updated Nov 6, 2020
Next