Skip to content
View hongpeng-guo's full-sized avatar
:octocat:
:octocat:

Block or report hongpeng-guo

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

This project aims to collect the latest "call for reviewers" links from various top CS/ML/AI conferences/journals

1,171 49 Updated Feb 6, 2026

[NSDI'26] PolyRL is a reinforcement learning framework for LLM that harvest spot instances on the cloud to reduce cost.

Python 19 1 Updated Mar 30, 2026

API for developing Balatro bots 🃏

Python 70 16 Updated Aug 9, 2026

Tile primitives for speedy kernels

Cuda 3,629 319 Updated Jul 13, 2026

On demand communication

Python 34 3 Updated Apr 16, 2026

Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞

TypeScript 386,071 81,141 Updated Aug 12, 2026

分享AI Infra知识&代码练习:PyTorch、vLLM/SGLang、slime/vime框架入门⚡️、性能加速🚀、大模型基础🧠、AI软硬件🔧等

Jupyter Notebook 3,480 329 Updated Aug 7, 2026

NexRL is an ultra-loosely-coupled LLM post-training framework.

Python 119 9 Updated Jul 30, 2026

A standard API for single-agent reinforcement learning environments, with popular reference environments and related utilities (formerly Gym)

Python 12,312 1,407 Updated Aug 5, 2026

A toolkit for developing and comparing reinforcement learning algorithms.

Python 37,242 8,684 Updated Mar 26, 2026

A NCCL extension library, designed to efficiently offload GPU memory allocated by the NCCL communication library.

C++ 116 11 Updated Dec 17, 2025

A compact implementation of SGLang, designed to demystify the complexities of modern LLM serving systems.

Python 4,739 785 Updated May 17, 2026
Python 1,053 101 Updated Jul 14, 2026

A set of examples based on verl for end-to-end RL training recipes.

Python 324 150 Updated Aug 5, 2026

Virtualized Elastic KV Cache for Dynamic GPU Sharing and Beyond

Python 1,131 130 Updated Aug 11, 2026

Training API and CLI

Python 691 78 Updated Aug 8, 2026

Long-RL: Scaling RL to Long Sequences (NeurIPS 2025)

Python 726 30 Updated Sep 24, 2025
Python 1,435 105 Updated Aug 6, 2026

Miles is an enterprise-facing reinforcement learning framework for LLM and VLM post-training, forked from and co-evolving with slime.

Python 1,973 354 Updated Aug 12, 2026

An Efficient and User-Friendly Scaling Library for Reinforcement Learning with Large Language Models

Python 3,360 305 Updated Aug 12, 2026

NVIDIA Inference Xfer Library (NIXL)

C++ 1,188 402 Updated Aug 12, 2026

Ideas for projects related to Tinker

195 12 Updated Nov 6, 2025

Local UI to run and train LLMs and diffusion models, including Qwen3.8, Kimi K3, MiniMax-H3, Gemma 4, DeepSeek-V4, FLUX and more.

Python 70,623 6,380 Updated Aug 12, 2026

Kimi K2 is the large language model series developed by Moonshot AI team

11,102 904 Updated Jan 21, 2026

A Lightweight LLM Post-Training Library

Python 2,401 330 Updated Aug 12, 2026

Virtual whiteboard for sketching hand-drawn like diagrams

TypeScript 129,443 14,833 Updated Aug 12, 2026

Mooncake is the serving platform for Kimi, a leading LLM service provided by Moonshot AI.

C++ 6,255 1,074 Updated Aug 12, 2026

Post-training with Tinker

Python 4,013 509 Updated Aug 12, 2026

Checkpoint-engine is a simple middleware to update model weights in LLM inference engines

Python 995 102 Updated Aug 12, 2026
Next