Skip to content
View JaejinCho's full-sized avatar

Block or report JaejinCho

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

A LLM-based Agent that predict its tasks proactively.

Python 646 64 Updated May 12, 2026

Fully open reproduction of DeepSeek-R1

Python 26,433 2,445 Updated Apr 2, 2026

Awesome speech/audio LLMs, representation learning, and codec models

1,245 74 Updated Jul 10, 2026

Python toolkit for speech processing

Python 72 22 Updated Aug 4, 2026

Take Bible Study notes easily in the popular note-taking app Obsidian, with automatic verse and reference suggestions.

TypeScript 339 81 Updated Aug 7, 2026

Interactive e-book for Python to C++ transition

Python 53 59 Updated Aug 11, 2026

A framework to enable a multimodal model to operate a computer.

Python 10,282 1,424 Updated Sep 19, 2025

A programming framework for agentic AI

Python 60,386 9,099 Updated Apr 15, 2026

Code, Dataset, and Pretrained Models for Audio and Speech Large Language Model "Listen, Think, and Understand".

Python 478 41 Updated Apr 24, 2024

šŸ¤— Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.

Python 164,001 34,223 Updated Aug 12, 2026

Foundational Models for State-of-the-Art Speech and Text Translation

Jupyter Notebook 11,839 1,176 Updated Jul 28, 2026

Audiocraft is a library for audio processing and generation with deep learning. It features the state-of-the-art EnCodec audio compressor / tokenizer, along with MusicGen, a simple and controllable…

Jupyter Notebook 23,559 2,679 Updated Mar 3, 2026

A scalable generative AI framework built for researchers and developers working on Large Language Models, Multimodal, and Speech AI (Automatic Speech Recognition and Text-to-Speech)

Python 18,107 3,550 Updated Aug 12, 2026

Awesome-LLM: a curated list of Large Language Model

27,255 2,672 Updated Jul 31, 2025

A playbook for systematically maximizing the performance of deep learning models.

30,279 2,420 Updated Jun 18, 2024

Personal homepage for Desh Raj

CSS 14 37 Updated Jan 12, 2026

A curated list of awesome Speaker Diarization papers, libraries, datasets, and other resources.

1,890 242 Updated Aug 12, 2026

Robust Speech Recognition via Large-Scale Weak Supervision

Python 107,146 13,002 Updated Jul 28, 2026

This repository contains demos I made with the Transformers library by HuggingFace.

Jupyter Notebook 11,741 1,734 Updated Apr 20, 2026

Think DSP: Digital Signal Processing in Python, by Allen B. Downey.

Jupyter Notebook 4,602 3,585 Updated Aug 8, 2026
Python 217 54 Updated Jan 23, 2026

Development Toolkit for the VoxCeleb Speaker Recognition Challenge 2021

Perl 19 Updated Jul 21, 2021

A curated list of awesome self-supervised methods

6,411 836 Updated Feb 24, 2026
Python 4 1 Updated May 8, 2020

Contrastive Predictive Coding for Automatic Speaker Verification

Python 506 97 Updated Oct 29, 2019

End-to-End Speech Processing Toolkit

Python 9,919 2,420 Updated Aug 11, 2026