Skip to content
View MaoZiming's full-sized avatar
🔭
Thinking
🔭
Thinking

Organizations

@NetSys @Y-Hack @Yale-LILY @ucbsky @yale-nova @skypilot-org @berkeley-cs168 @Trinity-data-store @uccl-project

Block or report MaoZiming

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Scalable toolkit for efficient model reinforcement

Python 1,903 512 Updated Aug 13, 2026

slime is an LLM post-training framework for RL Scaling.

Python 7,877 1,135 Updated Aug 12, 2026

Miles is an enterprise-facing reinforcement learning framework for LLM and VLM post-training, forked from and co-evolving with slime.

Python 1,977 356 Updated Aug 13, 2026

htop-like TUI for real-time RDMA network monitoring.

Rust 80 6 Updated Aug 11, 2026

Can LLMs Write Correct and Efficient GPU Communication Code?

Python 60 2 Updated Jul 7, 2026

The end of web parsing. The beginning of scalable pixel-native search. link: https://pixelrag.ai/

Python 9,486 813 Updated Jul 31, 2026

mKernel: fast multi-node, multi-GPU fused kernels

Cuda 264 24 Updated Aug 12, 2026

Ring attention implementation with flash attention

Python 1,045 101 Updated Sep 10, 2025

Fast and memory-efficient exact kmeans

Python 708 44 Updated Aug 4, 2026

A Proof-oriented Programming Language

F* 3,095 261 Updated Aug 12, 2026
Python 454 46 Updated Aug 12, 2026

A Distributed Attention Towards Linear Scalability for Ultra-Long Context, Heterogeneous Data Training

Python 908 65 Updated Aug 11, 2026

A benchmark of real-world DL kernel problems

Python 276 33 Updated Jul 15, 2026

Share your GPU without MIG or MPS

Python 49 4 Updated Jan 27, 2026

Google Workspace CLI — one command-line tool for Drive, Gmail, Calendar, Sheets, Docs, Chat, Admin, and more. Dynamically built from Google Discovery Service. Includes AI agent skills.

Rust 30,351 1,779 Updated Aug 1, 2026

Automated High-Performance GPU Kernel Generation

Python 129 26 Updated Jun 1, 2026

Fast and Furious AMD Kernels

C++ 454 78 Updated Aug 13, 2026

CUDA-L1: Improving CUDA Optimization via Contrastive Reinforcement Learning

Python 316 108 Updated Nov 3, 2025
Shell 24 3 Updated Jan 18, 2026

DeepGEMM: clean and efficient BLAS kernel library on GPU

Cuda 7,658 1,165 Updated Aug 11, 2026

Mirage Persistent Kernel: Compiling LLMs into a MegaKernel

Cuda 2,419 238 Updated Aug 4, 2026

Research works from Tencent AI Lab regarding self-evolving agents

Python 98 5 Updated Jan 30, 2026

SysMoBench: Evaluating AI on Formally Modeling Complex Real-World Systems

Python 23 3 Updated Jul 13, 2026

Autonomous GPU Kernel Generation & Optimization via Deep Agents

Python 509 86 Updated Jul 15, 2026

Building the Virtuous Cycle for AI-driven LLM Systems

Python 265 47 Updated May 1, 2026

A fast communication-overlapping library for tensor/expert parallelism on GPUs.

C++ 1,353 113 Updated Aug 28, 2025

Tile-based language built for AI computation across all scales

Python 185 9 Updated Aug 12, 2026

Tile-Based Runtime for Ultra-Low-Latency LLM Inference

Python 1,676 119 Updated Aug 13, 2026
Next