Skip to content
View shanguanma's full-sized avatar
🎯
Focusing
🎯
Focusing
  • The Chinese University of Hong Kong, Shenzhen(CUHK-SZ); Shenzhen Research Institute of Big Data(SRIBD)
  • Shenzhen

Block or report shanguanma

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

DeepSeek Harness: Everything is a Plugin.

TypeScript 175,858 19,056 Updated Aug 19, 2026

Generate production-quality SVG+PNG technical diagrams from natural language. 7 styles, UML support, and AI/Agent workflow patterns.

Python 10,719 870 Updated Aug 18, 2026

A high-throughput and memory-efficient inference and serving engine for LLMs

Python 89,585 20,978 Updated Aug 21, 2026

Practical quantization recipes for large language models and speech models, from model preparation through deployment-oriented validation.

Python 29 5 Updated Aug 14, 2026

Real-time text-to-speech with Qwen3-TTS

Python 1,320 191 Updated Jul 17, 2026

LLM model quantization (compression) toolkit with HW acceleration support for Nvidia, AMD, Intel GPU and Intel/AMD/Apple CPU via HF, vLLM, and SGLang.

Python 1,235 202 Updated Aug 19, 2026

精心收集的天涯神贴,不带水印,方便阅读

2,941 708 Updated Jul 28, 2026

Native End-to-End Full-Duplex Spoken Language Model

Python 121 7 Updated Aug 18, 2026

Codex skill for converting slide images, PDFs, and image-based PPTX files into editable PowerPoint decks.

Python 2,083 125 Updated Jul 28, 2026

AI agent toolkit: unified LLM API, agent loop, TUI, coding agent CLI

TypeScript 94,545 11,701 Updated Aug 20, 2026

Run frontier MoE models on hardware you already own — pure C, zero deps, experts streamed from disk. Tiny engine, immense model. 🐦

C 25,623 2,789 Updated Aug 20, 2026

GPT-2 training in pure Mojo with hand-written CUDA and Metal GPU kernels. llm.c parity in bf16 on NVIDIA, 1.72x faster than PyTorch MPS on Apple Silicon.

Mojo 23 5 Updated Aug 13, 2026

A new bootable USB solution.

C 78,849 4,898 Updated Aug 6, 2026

Framework for orchestrating role-playing, autonomous AI agents. By fostering collaborative intelligence, CrewAI empowers agents to work together seamlessly, tackling complex tasks.

Python 57,402 8,198 Updated Aug 21, 2026

A framework for efficient model inference with omni-modality models

Python 6,206 1,512 Updated Aug 21, 2026

Converts text to speech in realtime

Python 4,012 403 Updated Aug 21, 2026

10000 chatTTS voices !chatTTS 音色库,再也不为音色抽卡烦恼啦。这是我第一个项目,熬夜龟速生产10000条音色并上传Github,给点鼓励呗哈!主域名:https://www.TTSlist.com 备用:http://ttslist.aiqbh.com/

HTML 256 21 Updated Jul 18, 2024

Qwen3-ASR is an open-source series of ASR models developed by the Qwen team at Alibaba Cloud, supporting stable multilingual speech/music/song recognition, language detection and timestamp prediction.

Python 3,391 336 Updated Jun 26, 2026

Ultimate Vocal Remover Inference CLI

Python 123 12 Updated Feb 27, 2026

解决 Cursor 使用 DeepSeek V4 模型时的 `reasoning_content must be passed back` 错误

Python 147 14 Updated May 14, 2026
Python 33 5 Updated Aug 21, 2026

Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞

TypeScript 386,937 81,274 Updated Aug 21, 2026

Mirror of https://git.ffmpeg.org/ffmpeg.git

C 63,484 14,170 Updated Aug 21, 2026

A theoretical reconstruction of the Claude Mythos architecture, built from first principles using the available research literature.

Python 14,799 3,293 Updated May 23, 2026

Verbose implementations of LLMs architectures, techniques and research papers from scratch. Qwen..., RLHF, Multimodal, Hyperconnections...

Python 15 2 Updated Apr 11, 2026

FastAPI-compatible Python framework with Zig HTTP core; 7x faster, free-threading native

Zig 1,001 29 Updated Aug 17, 2026

Kyutai's Speech-To-Text and Text-To-Speech models based on the Delayed Streams Modeling framework.

Python 3,014 313 Updated Jan 26, 2026

AirLLM 70B inference with single 4GB GPU

Jupyter Notebook 31,962 3,385 Updated Aug 20, 2026

Train transformer language models with reinforcement learning.

Python 19,113 2,920 Updated Aug 21, 2026

A compact implementation of SGLang, designed to demystify the complexities of modern LLM serving systems.

Python 4,781 795 Updated May 17, 2026
Next