Highlights
- Pro
Lists (1)
Sort Name ascending (A-Z)
Stars
X-IL: Exploring the Design Space of Imitation Learning Policies
This project aims to collect the latest "call for reviewers" links from various top CS/ML/AI conferences/journals
Minimal reproduction of DeepSeek R1-Zero
[NeurIPS 2025 Spotlight] Reasoning Environments for Reinforcement Learning with Verifiable Rewards
A beautiful, simple, clean, and responsive Jekyll theme for academics
Low ReSource Reinforcement Learning with CPU Offloading Training Support
Minecraft AI with LLMs+Mineflayer
A benchmark environment for fully cooperative human-AI performance.
🤗 LeRobot: Making AI for Robotics more accessible with end-to-end learning
Build your own visual reasoning model
[AAAI 2026] - Official repo for paper: "Reinforcement Learning for Reasoning in Small LLMs: What Works and What Doesn't"
RENT (Reinforcement Learning via Entropy Minimization) is an unsupervised method for training reasoning LLMs.
SLIT-AI / ADPA
Forked from gaoshiping/ADPA[ICLR2025 Spotlight] Advantage-Guided Distillation for Preference Alignment in Small Language Models
This repository collects papers for "A Survey on Knowledge Distillation of Large Language Models". We break down KD into Knowledge Elicitation and Distillation Algorithms, and explore the Skill & V…
[TMLR 2025] Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models