Stars
Fast and memory-efficient exact attention
Open-source alternative to Claude Agent SDK, ChatGPT Agents, and Manus.
An Open-Source Large-Scale Reinforcement Learning Project for Search Agents
verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework
The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.
Super-Efficient RLHF Training of LLMs with Parameter Reallocation
A machine learning compiler for GPUs, CPUs, and ML accelerators
Zero Bubble Pipeline Parallelism
Large Language Model (LLM) Systems Paper List
[ECCV 2024, Oral] DynamiCrafter: Animating Open-domain Images with Video Diffusion Priors
Scalable toolkit for efficient model alignment
Open-Sora: Democratizing Efficient Video Production for All
Tile primitives for speedy kernels
SRL: Scaling Distributed Reinforcement Learning to Over Ten Thousand Cores
Knowledge Agents and Management in the Cloud
The Charm++ parallel programming system. Visit https://charmplusplus.org/ for more information.
A high-performance distributed deep learning system targeting large-scale and automated distributed training. If you have any interests, please visit/star/fork https://github.com/PKU-DAIR/Hetu
Large Language Model Text Generation Inference
Automatically Discovering Fast Parallelization Strategies for Distributed Deep Neural Network Training
An automatic evaluator for instruction-following language models. Human-validated, high-quality, cheap, and fast.
FlagAI (Fast LArge-scale General AI models) is a fast, easy-to-use and extensible toolkit for large-scale model.
A library for accelerating Transformer models on NVIDIA GPUs, including using 8-bit and 4-bit floating point (FP8 and FP4) precision on Hopper, Ada and Blackwell GPUs, to provide better performance…
A scalable generative AI framework built for researchers and developers working on Large Language Models, Multimodal, and Speech AI (Automatic Speech Recognition and Text-to-Speech)
A repo for distributed training of language models with Reinforcement Learning via Human Feedback (RLHF)
An open platform for training, serving, and evaluating large language models. Release repo for Vicuna and Chatbot Arena.
Safe RLHF: Constrained Value Alignment via Safe Reinforcement Learning from Human Feedback
CodeGeeX: An Open Multilingual Code Generation Model (KDD 2023)
Central place for the engineering/scaling WG: documentation, SLURM scripts and logs, compute environment and data.