Skip to content
View Parakrant's full-sized avatar
😅
Lets go
😅
Lets go
  • SoundLab, CityU
  • Hong Kong
  • 22:35 (UTC -12:00)

Highlights

  • Pro

Block or report Parakrant

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results
Jupyter Notebook 10 Updated Jun 18, 2025

🔥 LeetCode for PyTorch — practice implementing softmax, attention, GPT-2 and more from scratch with instant auto-grading. Jupyter-based, self-hosted or try online.

Jupyter Notebook 4,468 390 Updated Jul 31, 2026

The official repo of UL-UNAS, an ultra-lightweight SE model.

Python 196 29 Updated Jun 17, 2026
Jupyter Notebook 3 Updated Jan 29, 2026
Python 132 11 Updated Jul 23, 2026

Encode and decode audio samples to/from continuous and discrete compressed representations!

Python 121 6 Updated Nov 25, 2025

A tutorial for Speech Enhancement researchers and practitioners. The purpose of this repo is to organize the world’s resources for speech enhancement and make them universally accessible and useful.

MATLAB 833 152 Updated Dec 1, 2020

A unified evaluation suite for speech-to-text translation, covering SpeechLLMs, SFMs, and cascaded systems across diverse real-world speech phenomena.

Jupyter Notebook 32 4 Updated Aug 9, 2026

The repository provides code for running inference with the Meta Segment Anything Audio Model (SAM-Audio), links for downloading the trained model checkpoints, and example notebooks that show how t…

Python 3,592 325 Updated May 26, 2026

MT3: Multi-Task Multitrack Music Transcription

Python 1,738 223 Updated Jul 9, 2026

A Conversational Speech Generation Model

Python 14,717 1,478 Updated May 27, 2025

Keep track of big models in audio domain, including speech, singing, music etc.

517 33 Updated Jul 3, 2026
Python 226 25 Updated Dec 5, 2024

Style transfer of synthetic electric guitar to more realistic electric guitar audio using flow matching.

Jupyter Notebook 14 1 Updated Nov 18, 2025

An implementation of the Sottek Hearing Model psychoacoustic sound quality metrics defined in ECMA-418-2.

Python 23 2 Updated Jun 1, 2026

An enhanced version of All-In-One with integrated source separation and modern PyTorch compatibility

Python 23 2 Updated Jul 21, 2026

Compute distribution-based quality metrics for audio data using embeddings, with a focus on music.

Python 47 3 Updated Jan 15, 2026

Official repository for the paper "MambAttention: Mamba with Multi-Head Attention for Generalizable Single-Channel Speech Enhancement" (Accepted to IEEE Transactions on Audio, Speech, and Language …

Python 35 5 Updated Mar 25, 2026

Speed-optimized streaming neural speech enhancement network

Python 145 36 Updated Jul 3, 2026

A must-read paper for speech separation based on neural networks

TypeScript 954 141 Updated Aug 11, 2025
Python 65 4 Updated Oct 10, 2025

Official implementation of the Sheet Music Transformer

Python 80 27 Updated Mar 24, 2026

Code of our ISMIR 2025 paper - D. Afchar, G. Meseguer Brocal, K. Akesbi, R. Hennequin

Jupyter Notebook 48 7 Updated Nov 12, 2025

Fast linear discrete time filtering in PyTorch.

Python 33 2 Updated Aug 9, 2026

Official repository for the paper - SLAP: Siamese Language-Audio Pretraining without negative samples for Music Understanding

Python 63 2 Updated Sep 25, 2025
Python 19 2 Updated Sep 20, 2025

A deep learning model for dynamic range compression modeling. Accompanying repository of the DAFx 2025 paper "Empirical Results for Adjusting Truncated Backpropagation Through Time while Training N…

Python 10 Updated Apr 23, 2026

A Large Dataset of Paired Guitar Audio Recordings and Tablatures

Jupyter Notebook 26 Updated Sep 30, 2025

DDSP experiments in Faust

30 1 Updated Feb 12, 2025

(ICASSP 2025, official code)FlowSE: Flow Matching-based Speech Enhancement

Python 108 5 Updated Jul 23, 2025
Next