Skip to content
View ywh555hhh's full-sized avatar
😇
已取证
😇
已取证
  • CityU
  • china
  • 12:46 (UTC +08:00)

Organizations

@NCUSCC @hust-open-atom-club

Block or report ywh555hhh

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Starred repositories

Showing results

High-performance Inference and Deployment Toolkit for LLMs and VLMs based on PaddlePaddle

Python 3,719 756 Updated Aug 26, 2026

🥷 Engineering habits you already know, turned into skills Claude can run.

Python 7,084 416 Updated Sep 22, 2026

Let's Learn AI SYStem

Python 53 405 Updated Aug 17, 2026

A rewrite of Tachiyomi for the Desktop

Java 7,747 460 Updated Sep 23, 2026

SGLang is a high-performance serving framework for large language models and multimodal models.

Python 36,384 9,108 Updated Sep 24, 2026

Kernel Design Agents (KDA) is a agent-centric workflow to write high-performance CUDA Kernels.

1,086 99 Updated Sep 14, 2026
Python 5 Updated Apr 29, 2026
Python 42 24 Updated Sep 21, 2026

A tutorial on modern GPU programming for machine learning systems

HTML 1,293 139 Updated Sep 3, 2026

An AI-native distributed data plane built in Rust that supports high performance RPC, KV Cache, Message Queue, and File & Object Acceleration

Rust 126 6 Updated Aug 10, 2026

DeepGEMM: clean and efficient BLAS kernel library on GPU

Cuda 7,864 1,275 Updated Sep 23, 2026

Skills for Real Engineers. Straight from my .agents directory.

Shell 268,581 22,644 Updated Sep 18, 2026
Go 97 9 Updated Sep 15, 2025

花笺,轻量优雅的跨平台桌面便签工具,支持 Markdown 编辑与预览

Rust 5,280 305 Updated Sep 11, 2026

Sub2API 一站式开源中转服务,让 Claude、Openai 、Gemini、Grok订阅统一接入,支持拼车共享,更高效分摊成本,原生工具无缝使用。

Go 42,557 9,087 Updated Sep 23, 2026

PiKV: KV Cache Management System for Mixture of Experts [Efficient ML System]

Python 63 8 Updated Aug 17, 2026

SGLang-Omni is a high-performance serving framework for audio models (TTS, ASR) and unified multimodal models.

Python 1,267 521 Updated Sep 24, 2026

A framework for efficient model inference with omni-modality models

Python 7,041 1,781 Updated Sep 24, 2026

Mooncake is the serving platform for Kimi, a leading LLM service provided by Moonshot AI.

C++ 6,649 1,261 Updated Sep 24, 2026

High-performance RL post-training infrastructure. Designed to achieve bitwise operator-level train-inference consistency across heterogeneous engines and extreme memory efficiency for GRPO, PPO, etc.

Python 369 96 Updated Sep 23, 2026

Pure Rust + CUDA LLM inference engine — no PyTorch, OpenAI-compatible, serves Qwen3 to Kimi-K2

Rust 712 108 Updated Sep 23, 2026

An AI-powered IDE for long-form fiction writing, combining software engineering workflows, modern storytelling methodologies, and multi-agent systems.

TypeScript 692 66 Updated Sep 22, 2026

Compile SillyTavern character cards into pi-native, event-driven interactive narrative runtimes.

Python 118 7 Updated Jul 11, 2026

AI agent toolkit: unified LLM API, agent loop, TUI, coding agent CLI

TypeScript 108,973 13,841 Updated Sep 23, 2026
Next