Skip to content
View gitwukeyi's full-sized avatar

Block or report gitwukeyi

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

An end-to-end framework for joint speech enhancement (SE) and automatic gain control (AGC)

Python 22 3 Updated Jul 6, 2026
Python 94 5 Updated Jun 25, 2025
Python 18 7 Updated Mar 9, 2025

Production First and Production Ready End-to-End Keyword Spotting Toolkit

Python 748 147 Updated Jul 23, 2026

The official Pytorch implementation of "Frame-wise streaming end-to-end speaker diarization with non-autoregressive self-attention-based attractors". [ICASSP 2024] and "LS-EEND: long-form streaming…

Python 187 17 Updated May 7, 2026

The official repo of UL-UNAS, an ultra-lightweight SE model.

Python 196 29 Updated Jun 17, 2026

An unofficial implementation of DeepVQE proposed by Microsoft Corp.

Python 148 36 Updated Mar 24, 2025

partitioned block based frequency domain Kalman filter

Python 58 9 Updated Jan 14, 2023

The dataset of Speech Recognition

467 81 Updated Jan 4, 2026

Port of OpenAI's Whisper model in C/C++

C++ 52,806 6,055 Updated Aug 7, 2026

gpt-oss-120b and gpt-oss-20b are two open-weight language models by OpenAI

Python 20,304 2,133 Updated Jul 24, 2026

Open-source framework for conversational voice AI agents

Python 11,036 1,348 Updated Aug 6, 2026
Python 214 35 Updated Dec 4, 2023

Official repository for U-SAM (Interspeech 2025)

Python 28 4 Updated Jun 3, 2025

The official repo of NBC & SpatialNet for multichannel speech separation, denoising, and dereverberation

Python 365 45 Updated Jan 1, 2025

Voice Activity Detector (VAD) : low-latency, high-performance and lightweight

C 2,227 175 Updated Feb 2, 2026

SonicSim: A customizable simulation platform for speech processing in moving sound source scenarios

Python 279 23 Updated Jan 22, 2025

Muon is an optimizer for hidden layers in neural networks

Python 2,774 128 Updated May 24, 2026

An AI-Powered Speech Processing Toolkit and Open Source SOTA Pretrained Models, Supporting Speech Enhancement, Separation, and Target Speaker Extraction, etc.

Python 4,398 359 Updated Aug 14, 2025

百聆 是一个类似GPT-4o的语音对话机器人,通过ASR+LLM+TTS实现,集成DeepSeek R1等优秀大模型,接入openClaw,真正的个人语音助手,时延低至800ms,Mac等低配置也可运行,支持打断

Python 1,750 307 Updated Apr 6, 2026

Silero Models: pre-trained text-to-speech models made embarrassingly simple

Jupyter Notebook 6,060 370 Updated Jul 31, 2026

Collect super-resolution related papers, data, repositories

3,092 369 Updated Jun 11, 2026

This repository is the official implementation of the ICASSP 2025 paper: 'Self-supervised blind room parameter estimation using attention mechanisms.' The code will be made available upon the accep…

Jupyter Notebook 4 1 Updated May 7, 2025

Fully Quantized Neural Networks For Speech Enhancement

Python 65 12 Updated Feb 15, 2024

Acoustic Echo Canceller for Mobile Module Port From WebRTC

C 216 102 Updated Apr 28, 2025

A Flutter project can make you watch live with ease.

Dart 1,466 63 Updated Jan 10, 2024
Python 59 14 Updated Apr 24, 2024

校招、秋招、春招、实习好项目!带你从零实现一个高性能的深度学习推理库,支持大模型 llama2 、Unet、Yolov5、Resnet等模型的推理。Implement a high-performance deep learning inference library step by step

C++ 3,490 371 Updated Jun 22, 2025

Facebook AI Research Sequence-to-Sequence Toolkit written in Python.

Python 32,243 6,674 Updated Sep 30, 2025
Next