Skip to content
View zll961020's full-sized avatar

Block or report zll961020

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results
Python 1 Updated May 18, 2026

MOSS-Audio is an open-source foundation model for unified audio understanding, enabling speech, sound, music, captioning, QA, and reasoning in real-world scenarios.

Python 1 Updated Jun 2, 2026

First foundation ASR built for the real world - 7 atomic acoustic conditions, 54 compound scenarios, 2.6M samples, and up to ~30% gains over SOTA where every other model falls apart. **You'll come …

Python 1 Updated Jun 3, 2026

仅需Python基础,从0构建自己的具身智能机器人;从0逐步构建VLA/OpenVLA/SmolVLA/Pi0, 深入理解具身智能

Python 1 Updated Jul 2, 2026

A Large-scale Wu Dialect Speech Corpus with Multi-dimensional Annotations

Python 171 4 Updated Feb 6, 2026

A lightweight suite of motion imitation methods for training controllers.

Python 1 Updated Jun 23, 2026

Murmur: An Efficient Inference System for Long-Form ASR

Python 9 2 Updated Jun 3, 2026

Official repository for the WenetSpeech-Chuan dataset.

Python 217 6 Updated Jul 14, 2026

A Code Release for Mip-NeRF 360, Ref-NeRF, and RawNeRF

Python 1 Updated Dec 8, 2023

Contrastive Language-Audio Pretraining

Python 1 Updated May 15, 2025

📚 《从零开始构建智能体》——从零开始的智能体原理与实践教程

Python 1 Updated Jan 4, 2026

Open-Source Frontier Voice AI

Python 1 Updated Jan 27, 2026

Qwen3-ASR is an open-source series of ASR models developed by the Qwen team at Alibaba Cloud, supporting stable multilingual speech/music/song recognition, language detection and timestamp prediction.

Python 1 Updated Jan 30, 2026

A toolkit for speaker diarization.

Jupyter Notebook 1 Updated Jan 31, 2026
Python 1 Updated Feb 3, 2026

Claude Code v2.1.88 Source Code

TypeScript 1 Updated Mar 31, 2026

A visual, example-driven guide to Claude Code — from basic concepts to advanced agents, with copy-paste templates that bring immediate value.

Python 1 Updated Apr 11, 2026

Agent harness built with LangChain and LangGraph. Equipped with a planning tool, a filesystem backend, and the ability to spawn subagents - well-equipped to handle complex agentic tasks.

Python 1 Updated Apr 15, 2026

An open-source long-horizon SuperAgent harness that researches, codes, and creates. With the help of sandboxes, memories, tools, skill, subagents and message gateway, it handles different levels of…

Python 1 Updated Apr 14, 2026

来自于文章Paraformer-v2: An improved non-autoregressive transformer for noise-robust speech recognition

Python 29 4 Updated Nov 20, 2024

Variational Bayes HMM over x-vectors diarization

Python 287 58 Updated Jan 15, 2024

An Efficient and User-Friendly Scaling Library for Reinforcement Learning with Large Language Models

Python 1 Updated Dec 22, 2025

Open-source LLM-based ASR model family for Chinese, dialect, accent, and multilingual speech, with FunASR, vLLM, streaming, and llama.cpp runtimes.

C 1,429 141 Updated Jul 24, 2026

🤗 R1-AQA Model: mispeech/r1-aqa

Python 1 Updated Mar 28, 2025

SALMONN family: A suite of advanced multi-modal LLMs

1 Updated Sep 28, 2025

Qwen3-omni is a natively end-to-end, omni-modal LLM developed by the Qwen team at Alibaba Cloud, capable of understanding text, audio, images, and video, as well as generating speech in real time.

Jupyter Notebook 1 Updated Oct 9, 2025

Official code release for "Efficient Perspective-Correct 3D Gaussian Splatting Using Hybrid Transparency"

Cuda 1 Updated Oct 16, 2025
Next