Lists (1)
Sort Name ascending (A-Z)
Stars
IronClaw is an Agent OS focused on privacy, security and extensibility
An Open Phone Agent Model & Framework. Unlocking the AI Phone for Everyone
STEP-GUI: The top GUI agent solution in the galaxy. Developed by the StepFun-GELab team and powered by StepFun’s cutting-edge research capabilities.
Official Repository of "Learning to Reason under Off-Policy Guidance"
GLM-4.6V/4.5V/4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning
Trinity-RFT is a general-purpose, flexible and scalable framework designed for reinforcement fine-tuning (RFT) of large language models (LLM).
[ICLR 2026] On the Generalization of SFT: A Reinforcement Learning Perspective with Reward Rectification.
Hierarchical Reasoning Model Official Release
A high-throughput and memory-efficient inference and serving engine for LLMs
SGLang is a high-performance serving framework for large language models and multimodal models.
Use PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs (Qwen3.6, DeepSeek-V4, GLM-5.1, InternLM3, Llama4, ...) and 300+ MLLMs (Qwen3-VL, Qwen3-Omni, InternVL3.5, Ovis2.5, GLM4.5v, Gemma4, Llava, …
Community maintained hardware plugin for vLLM on Ascend
Ongoing research training transformer models at scale
Making large AI models cheaper, faster and more accessible
Mirror for HUAWEI MindSpeed repository (https://gitee.com/ascend/MindSpeed)
Minimal reproduction of DeepSeek R1-Zero
OpenCompass is an LLM evaluation platform, supporting a wide range of models (Llama3, Mistral, InternLM2,GPT-4,LLaMa2, Qwen,GLM, Claude, etc) over 100+ datasets.
Recipes to scale inference-time compute of open models
A collection of LLM papers, blogs, and projects, with a focus on OpenAI o1 🍓 and reasoning techniques.
Simple speculative decoding technique, integrated in vLLM and transformers
Official Implementation of EAGLE-1 (ICML'24), EAGLE-2 (EMNLP'24), and EAGLE-3 (NeurIPS'25).
Medusa: Simple Framework for Accelerating LLM Generation with Multiple Decoding Heads
Mastery of a Three-Word Language for Knowledge Graph Completion