Skip to content
View confiself's full-sized avatar

Block or report confiself

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

IronClaw is an Agent OS focused on privacy, security and extensibility

Rust 12,596 1,485 Updated Aug 10, 2026

An Open Phone Agent Model & Framework. Unlocking the AI Phone for Everyone

Python 25,976 4,016 Updated Mar 6, 2026

STEP-GUI: The top GUI agent solution in the galaxy. Developed by the StepFun-GELab team and powered by StepFun’s cutting-edge research capabilities.

Python 2,249 198 Updated May 11, 2026

Official Repository of "Learning to Reason under Off-Policy Guidance"

Python 461 72 Updated Mar 20, 2026

GLM-4.6V/4.5V/4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Python 2,363 176 Updated Jul 21, 2026

Trinity-RFT is a general-purpose, flexible and scalable framework designed for reinforcement fine-tuning (RFT) of large language models (LLM).

Python 682 75 Updated Jul 31, 2026

将SmolVLM2的视觉头与Qwen3-0.6B模型进行了拼接微调

Python 606 55 Updated Sep 8, 2025

[ICLR 2026] On the Generalization of SFT: A Reinforcement Learning Perspective with Reward Rectification.

Python 1,095 26 Updated Aug 1, 2026

Hierarchical Reasoning Model Official Release

Python 12,615 1,827 Updated Mar 31, 2026

A high-throughput and memory-efficient inference and serving engine for LLMs

Python 88,638 20,496 Updated Aug 10, 2026

SGLang is a high-performance serving framework for large language models and multimodal models.

Python 31,608 7,773 Updated Aug 10, 2026

Use PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs (Qwen3.6, DeepSeek-V4, GLM-5.1, InternLM3, Llama4, ...) and 300+ MLLMs (Qwen3-VL, Qwen3-Omni, InternVL3.5, Ovis2.5, GLM4.5v, Gemma4, Llava, …

Python 15,100 1,584 Updated Aug 9, 2026

Community maintained hardware plugin for vLLM on Ascend

C++ 2,593 2,010 Updated Aug 10, 2026
Python 228 38 Updated Jan 5, 2026

Ongoing research training transformer models at scale

Python 17,384 4,346 Updated Aug 10, 2026

Making large AI models cheaper, faster and more accessible

Python 41,430 4,506 Updated Jul 13, 2026

Mirror for HUAWEI MindSpeed repository (https://gitee.com/ascend/MindSpeed)

Python 2 1 Updated May 24, 2026

Minimal reproduction of DeepSeek R1-Zero

Python 13,222 1,580 Updated Feb 27, 2026

Simple RL training for reasoning

Python 3,872 285 Updated Dec 23, 2025

OpenCompass is an LLM evaluation platform, supporting a wide range of models (Llama3, Mistral, InternLM2,GPT-4,LLaMa2, Qwen,GLM, Claude, etc) over 100+ datasets.

Python 7,288 830 Updated Aug 4, 2026

Recipes to scale inference-time compute of open models

Python 1,132 131 Updated May 26, 2026

A collection of LLM papers, blogs, and projects, with a focus on OpenAI o1 🍓 and reasoning techniques.

6,897 369 Updated Dec 17, 2025
Python 972 110 Updated Jan 23, 2025

Simple speculative decoding technique, integrated in vLLM and transformers

Jupyter Notebook 612 28 Updated Aug 23, 2024

Official Implementation of EAGLE-1 (ICML'24), EAGLE-2 (EMNLP'24), and EAGLE-3 (NeurIPS'25).

Python 2,499 297 Updated Feb 20, 2026

Medusa: Simple Framework for Accelerating LLM Generation with Multiple Decoding Heads

Jupyter Notebook 2,764 205 Updated Jun 25, 2024

Mastery of a Three-Word Language for Knowledge Graph Completion

Python 33 3 Updated Feb 24, 2025

SimBERT升级版(SimBERTv2)!

Python 441 75 Updated Mar 21, 2022
Next