Skip to content
View hongpeng-guo's full-sized avatar
:octocat:
:octocat:

Block or report hongpeng-guo

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

This project aims to collect the latest "call for reviewers" links from various top CS/ML/AI conferences/journals

1,172 49 Updated Feb 6, 2026

[NSDI'26] PolyRL is a reinforcement learning framework for LLM that harvest spot instances on the cloud to reduce cost.

Python 19 1 Updated Mar 30, 2026

API for developing Balatro bots 🃏

Python 71 16 Updated Aug 9, 2026

Tile primitives for speedy kernels

Cuda 3,632 319 Updated Jul 13, 2026

On demand communication

Python 33 3 Updated Apr 16, 2026

Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞

TypeScript 386,396 81,209 Updated Aug 15, 2026

分享AI Infra知识&代码练习:PyTorch、vLLM/SGLang、slime/vime框架入门⚡️、性能加速🚀、大模型基础🧠、AI软硬件🔧等

Jupyter Notebook 3,522 335 Updated Aug 7, 2026

NexRL is an ultra-loosely-coupled LLM post-training framework.

Python 118 9 Updated Jul 30, 2026

A standard API for single-agent reinforcement learning environments, with popular reference environments and related utilities (formerly Gym)

Python 12,325 1,410 Updated Aug 5, 2026

A toolkit for developing and comparing reinforcement learning algorithms.

Python 37,239 8,685 Updated Mar 26, 2026

A NCCL extension library, designed to efficiently offload GPU memory allocated by the NCCL communication library.

C++ 115 11 Updated Dec 17, 2025

A compact implementation of SGLang, designed to demystify the complexities of modern LLM serving systems.

Python 4,745 787 Updated May 17, 2026
Python 1,054 103 Updated Jul 14, 2026

A set of examples based on verl for end-to-end RL training recipes.

Python 324 152 Updated Aug 5, 2026

Virtualized Elastic KV Cache for Dynamic GPU Sharing and Beyond

Python 1,134 131 Updated Aug 13, 2026

Training API and CLI

Python 695 78 Updated Aug 8, 2026

Long-RL: Scaling RL to Long Sequences (NeurIPS 2025)

Python 726 30 Updated Sep 24, 2025
Python 1,436 105 Updated Aug 6, 2026

Miles is an enterprise-facing reinforcement learning framework for LLM and VLM post-training, forked from and co-evolving with slime.

Python 1,999 360 Updated Aug 15, 2026

An Efficient and User-Friendly Scaling Library for Reinforcement Learning with Large Language Models

Python 3,362 305 Updated Aug 15, 2026

NVIDIA Inference Xfer Library (NIXL)

C++ 1,193 404 Updated Aug 15, 2026

Ideas for projects related to Tinker

196 11 Updated Nov 6, 2025

Local UI to run and train LLMs and diffusion models, including Qwen3.8, Kimi K3, MiniMax-H3, Gemma 4, DeepSeek-V4, FLUX and more.

Python 71,995 6,488 Updated Aug 15, 2026

Kimi K2 is the large language model series developed by Moonshot AI team

11,105 905 Updated Jan 21, 2026

A Lightweight LLM Post-Training Library

Python 2,406 332 Updated Aug 15, 2026

Virtual whiteboard for sketching hand-drawn like diagrams

TypeScript 129,667 14,892 Updated Aug 15, 2026

Mooncake is the serving platform for Kimi, a leading LLM service provided by Moonshot AI.

C++ 6,283 1,085 Updated Aug 15, 2026

Post-training with Tinker

Python 4,020 512 Updated Aug 15, 2026

Checkpoint-engine is a simple middleware to update model weights in LLM inference engines

Python 996 102 Updated Aug 12, 2026
Next