Skip to content
View light1726's full-sized avatar
🏠
Working from home
🏠
Working from home

Highlights

  • Pro

Organizations

@thuhcsi

Block or report light1726

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Repository of Streaming LLMs

Python 94 7 Updated Jul 26, 2026

DuplexSLA: A Full-Duplex Spoken Language Model with Synchronized Speech, Language, and Action

119 Updated May 20, 2026

Skills for Real Engineers. Straight from my .agents directory.

Shell 207,815 17,942 Updated Aug 6, 2026

The best-benchmarked open-source AI memory system. And it's free.

Python 58,165 7,480 Updated Aug 7, 2026

An agent-managed museum exhibit, built in Rust with Gajae-Code / LazyCodex — developed and maintained with no human intervention.

Rust 195,002 109,274 Updated Aug 6, 2026

Comprehensive open-source library of AI research and engineering skills for any AI model. Package the skills and your claude code/codex/gemini agent will be an AI research agent with full horsepowe…

TeX 11,482 840 Updated Jun 16, 2026

This is an evolving repo for the paper “From Turn-Taking to Synchronous Dialogue: A Survey of Full-Duplex Spoken Language Models ”A comprehensive survey of Full-Duplex Spoken Language Models (FD-SL…

29 Updated Dec 23, 2025

MichiAI: A Low Latency, Full Duplex Speech LLM with zero coherence loss

112 7 Updated Apr 24, 2026

A Pocket-Sized MLLM for Ultra-Efficient Image and Video Understanding on Your Phone

Python 26,113 2,040 Updated Aug 4, 2026

Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞

TypeScript 385,427 81,020 Updated Aug 7, 2026

Fun-Audio-Chat is a Large Audio Language Model built for natural, low-latency voice interactions.

Python 990 106 Updated Feb 27, 2026

The best ChatGPT that $100 can buy.

Python 57,030 7,906 Updated Aug 2, 2026

Step-Audio 2 is an end-to-end multi-modal large language model designed for industry-strength audio understanding and speech conversation.

Python 1,490 111 Updated Mar 16, 2026

A native-PyTorch library for large scale M-LLM (text/audio) training with tp/cp/dp.

Python 233 30 Updated Jul 2, 2026
Python 69 8 Updated Sep 3, 2024

Yin pitch estimator in PyTorch

Python 119 6 Updated Nov 7, 2022

SOTA Open Source TTS

Python 32,073 2,755 Updated Aug 3, 2026

✨✨Latest Advances on Multimodal Large Language Models

17,971 1,132 Updated Aug 3, 2026

A generative speech model for daily dialogue.

Python 39,748 4,255 Updated Apr 10, 2026

A casual and simple ChatGPT Python script that can run using terminal (as long as you have an API). Support Azure API.

Python 20 2 Updated May 3, 2025

The official repository of Dynamic-SUPERB.

Python 200 90 Updated Jun 24, 2025

Easy-to-Use Speech MOS predictors

Python 364 18 Updated Oct 24, 2023

TTS FrontEnd DataSet: Polyphone / Prosody / TextNormalization

Python 105 15 Updated Feb 5, 2024

An official reimplementation of the method described in the INTERSPEECH 2021 paper - Speech Resynthesis from Discrete Disentangled Self-Supervised Representations.

Python 416 59 Updated Aug 29, 2023

SLMTokBench for paper "SpeechTokenizer: Unified Speech Tokenizer for Speech Large Language Models"

37 Updated Aug 29, 2023

FAIR Sequence Modeling Toolkit 2

Python 1,143 146 Updated Aug 3, 2026

A library that contains a rich collection of performant PyTorch model metrics, a simple interface to create new metrics, a toolkit to facilitate metric computation in distributed training and tools…

Python 247 59 Updated Jul 30, 2026

Inference Llama 2 in one file of pure C

C 19,942 2,602 Updated Aug 6, 2024
Next