Skip to content
View catqaq's full-sized avatar

Block or report catqaq

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

A summary of technical reports for various large language models (LLMs).

Python 29 2 Updated Jun 22, 2026

Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞

TypeScript 383,959 80,664 Updated Jul 24, 2026

OpenClaw-RL: Train any agent simply by talking

Python 5,603 607 Updated May 23, 2026

Curated, opinionated index of post-R1 LLM × Reinforcement Learning. Many deep-dive blog posts cross-linked to many papers — GRPO, DAPO, DPO, PPO, RLHF, GSPO, CISPO, VAPO, Reward Modeling, MoE RL st…

71 6 Updated Jul 13, 2026

📖 This is a repository for organizing papers, codes, and other resources related to Latent Reasoning.

405 9 Updated Nov 5, 2025

High-performance Inference and Deployment Toolkit for LLMs and VLMs based on PaddlePaddle

Python 3,703 754 Updated Jul 21, 2026

OpenLLMAI-Research

2 Updated Jul 6, 2026

slime is an LLM post-training framework for RL Scaling.

Python 7,611 1,092 Updated Jul 24, 2026

SGLang is a high-performance serving framework for large language models and multimodal models.

Python 30,684 7,363 Updated Jul 24, 2026

verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework

Python 22,633 4,270 Updated Jul 24, 2026

Ongoing research training transformer language models at scale, including: BERT & GPT-2

Python 2,257 368 Updated Aug 14, 2025

Use PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs (Qwen3.6, DeepSeek-V4, GLM-5.1, InternLM3, Llama4, ...) and 300+ MLLMs (Qwen3-VL, Qwen3-Omni, InternVL3.5, Ovis2.5, GLM4.5v, Gemma4, Llava, …

Python 14,921 1,549 Updated Jul 22, 2026

Ongoing research training transformer models at scale

Python 17,188 4,277 Updated Jul 24, 2026

Lime: Explaining the predictions of any machine learning classifier

JavaScript 12,158 1,847 Updated Jul 25, 2024

AISystem 主要是指AI系统,包括AI芯片、AI编译器、AI推理和训练框架等AI全栈底层技术

Jupyter Notebook 17,259 2,423 Updated Sep 3, 2025
Python 1,341 54 Updated Nov 21, 2024

[ICML 2024] LESS: Selecting Influential Data for Targeted Instruction Tuning

Jupyter Notebook 532 47 Updated Oct 20, 2024

开源SFT数据集整理,随时补充

583 41 Updated Jun 2, 2023
Python 153 14 Updated Apr 16, 2024

Multipack distributed sampler for fast padding-free training of LLMs

Python 207 17 Updated Aug 10, 2024

Code for evaluating with Flow-Judge-v0.1 - an open-source, lightweight (3.8B) language model optimized for LLM system evaluations. Crafted for accuracy, speed, and customization.

Python 86 13 Updated Oct 29, 2024

Summarize existing representative LLMs text datasets.

1,478 149 Updated Mar 11, 2026

This is the first released survey paper on hallucinations of large vision-language models (LVLMs). To keep track of this field and continuously update our survey, we maintain this repository of rel…

96 7 Updated Jul 26, 2024

A Survey of LLM Alignment (SFT & RLHF), and A Survey of RLHF methods (2023~2024)

21 1 Updated May 21, 2024

Unsloth is a local UI for training and running Gemma 4, Qwen3.6, DeepSeek, Kimi, GLM and other models.

Python 68,807 6,190 Updated Jul 24, 2026
Python 336 32 Updated Jul 25, 2024

基于ClipCap的看图说话Image Caption模型

Python 325 44 Updated Apr 1, 2022

Similarities: a toolkit for similarity calculation and semantic search. 相似度计算、匹配搜索工具包,支持亿级数据文搜文、文搜图、图搜图,python3开发,开箱即用。

Python 903 87 Updated Mar 5, 2026

Welcome to the Llama Cookbook! This is your go to guide for Building with Llama: Getting started with Inference, Fine-Tuning, RAG. We also show you how to solve end to end problems using Llama mode…

Jupyter Notebook 18,538 2,766 Updated May 19, 2026

Evaluate your LLM's response with Prometheus and GPT4 💯

Python 1,102 68 Updated Apr 25, 2025
Next