Stars
Implementation of Poly-attention, a higher-order self-attention proposed by Chakrabarti et al. of Columbia
A collection of weight space learning including papers, codes, and datasets.
Awesome papers on weight-space learning
[ICLR 2025] NeuroLM: A Universal Multi-task Foundation Model for Bridging the Gap between Language and EEG Signals
Stable and Efficient Reinforcement Learning for Trillion-Parameter LLMs
Official repository of PEFT-Arena: Understanding Parameter-Efficient Finetuning from a Stability-Plasticity Perspective
The Cloud Sandbox Built for AI Agents
Transform geospatial relations into graphs for Graph Neural Networks and spatial network analysis
A clean implementation based on AlphaZero for any game in any framework + tutorial + Othello/Gobang/TicTacToe/Connect4 and more
A Survey of Reinforcement Learning for Large Reasoning Models
A curated list of awesome exploration RL resources (continually updated)
Muon is an optimizer for hidden layers in neural networks
A MemAgent framework that can be extrapolated to 3.5M, along with a training framework for RL training of any agent workflow.
Awesome In-Context RL: A curated list of In-Context Reinforcement Learning - - —
[NeurIPS 2025 Spotlight] Reasoning Environments for Reinforcement Learning with Verifiable Rewards
Understanding R1-Zero-Like Training: A Critical Perspective
Synthesizing Graphics Programs for Scientific Figures and Sketches with TikZ.
Fully open reproduction of DeepSeek-R1
Official Repo for Open-Reasoner-Zero
Learning Formal Mathematics from Intrinsic Motivation
An environment for learning formal mathematical reasoning from scratch
NOrangeeroli / deepclaude
Forked from winfunc/deepreasoningA high-performance LLM inference API and Chat UI that integrates DeepSeek R1's CoT reasoning traces with Anthropic Claude models.
A high-performance LLM inference API and Chat UI that integrates DeepSeek R1's CoT reasoning traces with Anthropic Claude models.