Skip to content
View xgwang's full-sized avatar

Block or report xgwang

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

TokenSpeed is a speed-of-light LLM inference engine.

Python 1,661 199 Updated Jul 26, 2026

GPU & Accelerator process monitoring for AMD, Apple, Huawei, Intel, NVIDIA and Qualcomm

C 10,863 412 Updated May 6, 2026

Secure and fast microVMs for serverless computing.

Rust 35,666 2,527 Updated Jul 24, 2026

Skills for Real Engineers. Straight from my .agents directory.

Shell 189,194 16,246 Updated Jul 23, 2026

The agent that grows with you

Python 220,775 42,062 Updated Jul 26, 2026

Open Multi-Agent Interactive Classroom — Get an immersive, multi-agent learning experience in just one click

TypeScript 20,211 3,978 Updated Jul 26, 2026

AI agents running research on single-GPU nanochat training automatically

Python 92,076 13,151 Updated Mar 26, 2026

Google Workspace CLI — one command-line tool for Drive, Gmail, Calendar, Sheets, Docs, Chat, Admin, and more. Dynamically built from Google Discovery Service. Includes AI agent skills.

Rust 30,013 1,753 Updated Jul 22, 2026

A request monitoring system for Claude Code that captures and visualizes all API requests and responses in real time. Helps developers monitor their Context for reviewing and debugging during Vibe …

JavaScript 1,045 122 Updated Jul 26, 2026

Browser automation CLI for AI agents

Rust 39,234 2,568 Updated Jul 26, 2026

Persistent file-based planning for AI coding agents and long-running tasks. Crash-proof markdown plans, session recovery after /clear and compaction, per-turn re-injection against context rot, dete…

Python 25,748 2,164 Updated Jul 24, 2026

A Flexible Framework for Experiencing Heterogeneous LLM Inference/Fine-tune Optimizations

Python 19,029 1,483 Updated Jul 24, 2026

Minimalistic 4D-parallelism distributed training framework for education purpose

Python 2,258 197 Updated Aug 26, 2025

LeetCode题解,151道题完整版。

TeX 11,340 3,374 Updated Jul 10, 2024

📚 技术面试必备基础知识、Leetcode、计算机操作系统、计算机网络、系统设计

184,895 50,828 Updated Aug 21, 2024

收集所有区块链(BlockChain)技术开发相关资料,包括Fabric和Ethereum开发资料

JavaScript 18,959 3,622 Updated Feb 29, 2024

Course to get into Large Language Models (LLMs) with roadmaps and Colab notebooks.

81,247 9,473 Updated Feb 5, 2026

A fast communication-overlapping library for tensor/expert parallelism on GPUs.

C++ 1,345 110 Updated Aug 28, 2025

Must read research papers and links to tools and datasets that are related to using machine learning for compilers and systems optimisation

1,680 178 Updated Jan 21, 2026

A fast GPU memory copy library based on NVIDIA GPUDirect RDMA technology

C 1,401 193 Updated Jul 14, 2026

Running large language models on a single GPU for throughput-oriented scenarios.

Python 9,363 590 Updated Oct 28, 2024

C++ standard library reference

HTML 839 60 Updated Dec 9, 2025

Probably the fastest coroutine lib in the world!

C++ 1,219 177 Updated Jul 24, 2026

CodeQL: the libraries and queries that power security researchers around the world, as well as code scanning in GitHub Advanced Security

CodeQL 9,869 2,030 Updated Jul 24, 2026

天涯 kkndme 神贴聊房价

19,421 3,852 Updated Jun 4, 2026

Confidential AI deployment with secure enclaves 🔒

Rust 512 36 Updated Mar 19, 2024

The LLVM Project is a collection of modular and reusable compiler and toolchain technologies.

LLVM 39,476 17,982 Updated Jul 26, 2026

Never ever ever use pixelation as a redaction technique

TypeScript 8,363 808 Updated Mar 15, 2024

The Triton Inference Server provides an optimized cloud and edge inferencing solution.

Python 10,869 1,820 Updated Jul 25, 2026
Next