Skip to content
View hw-han's full-sized avatar

Block or report hw-han

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

A lightning fast audio upsampler.

Python 776 72 Updated Feb 26, 2026

Open-Source Toolkit for End-to-End Speech Recognition leveraging PyTorch-Lightning and Hydra.

Python 715 115 Updated Jun 21, 2026

Noise supression using deep filtering

Python 4,601 500 Updated Oct 17, 2024

A comprehensive audio, image, video, CSV, and JSONL viewer extension for VSCode and Cursor.

TypeScript 35 5 Updated Aug 13, 2026

Open-Source Turn-Taking Detection Model and Dataset for Full-Duplex Spoken Dialogue Systems

Python 133 8 Updated Jan 25, 2026

Some comprehensive papers about speaker diarization

369 15 Updated Mar 24, 2026

Python Kalman filtering and optimal estimation library. Implements Kalman filter, particle filter, Extended Kalman filter, Unscented Kalman filter, g-h (alpha-beta), least squares, H Infinity, smoo…

Python 3,861 678 Updated Feb 7, 2024

Unofficial implementation of Titans, SOTA memory for transformers, in Pytorch

Python 1,975 207 Updated Jul 13, 2026

Official repository of SepReformer for speech separation

Python 263 42 Updated May 14, 2026

The official Pytorch implementation of "Frame-wise streaming end-to-end speaker diarization with non-autoregressive self-attention-based attractors". [ICASSP 2024] and "LS-EEND: long-form streaming…

Python 187 17 Updated May 7, 2026

Official implementation for our paper "Audio Mamba: Selective State Spaces for Self-Supervised Audio Representations"

Python 43 2 Updated Aug 14, 2025

wsj0-{2, 3, 4, 5} mix generation scripts, in Python.

Python 79 9 Updated Mar 17, 2021

Explicit Estimation of Magnitude and Phase Spectra in Parallel for High-Quality Speech Enhancement

Python 498 78 Updated May 19, 2025

Kalman Filter in Python (파이썬으로 구현하는 칼만 필터)

Jupyter Notebook 173 74 Updated Mar 13, 2020
Jupyter Notebook 7 2 Updated Jul 6, 2021

Collection of papers on state-space models

621 21 Updated Aug 10, 2026

Two-stage progressive neural network for acoustic echo cancellation

Python 56 12 Updated May 22, 2023

Audio processing project

Python 6 Updated Aug 9, 2022

Unofficial implementation of SCP-GAN

Python 18 1 Updated Jul 4, 2023

The official implementation of GTCRN, an ultra-lightweight SE model.

Python 714 116 Updated Aug 3, 2026

NVIDIA® TensorRT™ is an SDK for high-performance deep learning inference on NVIDIA GPUs. This repository contains the open source components of TensorRT.

C++ 13,255 2,391 Updated Aug 4, 2026

Kolmogorov Arnold Networks

Jupyter Notebook 16,334 1,561 Updated Jan 19, 2025

Variations of Kolmogorov-Arnold Networks

Python 116 11 Updated May 15, 2024

21 Lessons, Get Started Building with Generative AI

Jupyter Notebook 117,880 62,195 Updated Aug 13, 2026

This repo contains the official PyTorch implementation of "A Systematic Comparison of Phonetic Aware Techniques for Speech Enhancement" (Interspeech 2022)

Python 28 2 Updated Aug 8, 2022

Implementation of Mega, the Single-head Attention with Multi-headed EMA architecture that currently holds SOTA on Long Range Arena

Python 207 11 Updated Aug 26, 2023

A simple way to keep track of an Exponential Moving Average (EMA) version of your Pytorch model

Python 660 42 Updated Jul 31, 2026

ADAPTING SELF-SUPERVISED MODELS TO MULTI-TALKER SPEECH RECOGNITION USING SPEAKER EMBEDDINGS

Shell 34 1 Updated Mar 16, 2023

Large-scale Self-supervised Pre-training Across Tasks, Languages, and Modalities

Python 22,189 2,704 Updated Jan 23, 2026
Python 43 9 Updated May 27, 2024
Next