Skip to content
View pyemma's full-sized avatar

Block or report pyemma

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Starred repositories

Showing results

🤖 The analysis of Claude Code

TypeScript 3,810 2,040 Updated Apr 2, 2026

🧠「大模型」2小时完全从0训练64M的小参数LLM!Train a 64M-parameter LLM from scratch in just 2h!

Python 54,797 7,160 Updated Aug 6, 2026

An Efficient and User-Friendly Scaling Library for Reinforcement Learning with Large Language Models

Python 3,364 306 Updated Aug 18, 2026

🤗 Diffusers: State-of-the-art diffusion models for image, video, and audio generation in PyTorch.

Python 34,336 7,249 Updated Aug 18, 2026

Research on Coding Agents

12,224 19,586 Updated Apr 1, 2026

An educational resource to help anyone learn deep reinforcement learning.

Python 11,900 2,467 Updated Aug 5, 2024

The corresponding codes and dataset for OneSearch series

Python 168 18 Updated May 1, 2026

OpenClaw-RL: Train any agent simply by talking

Python 5,640 609 Updated May 23, 2026

A Claude Code skill that acts as your daily 军师 (strategic research advisor).

Shell 108 6 Updated Mar 18, 2026
Python 1,305 135 Updated May 20, 2026

An Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & VLM & TIS & vLLM & Ray & Async RL)

Python 9,930 1,002 Updated Aug 13, 2026

Implement a reasoning LLM in PyTorch from scratch, step by step

Jupyter Notebook 5,004 762 Updated Aug 4, 2026

EasyR1: An Efficient, Scalable, Multi-Modality RL Training Framework based on veRL

Python 5,122 385 Updated Jul 30, 2026

AI agents running research on single-GPU nanochat training automatically

Python 94,050 13,323 Updated Mar 26, 2026

Open-source RL Framework with Online Teacher-Student Distillation

Python 22 1 Updated Mar 5, 2026

Pure Triton kernels for Qwen3.5-27B inference on NVIDIA B200

Python 120 10 Updated Feb 28, 2026

Sparse Transition Matrix-Accelerated Trie Index for Constrained Decoding (https://arxiv.org/abs/2602.22647)

Python 230 28 Updated Mar 29, 2026

Minimalistic 4D-parallelism distributed training framework for education purpose

Python 2,282 200 Updated Aug 26, 2025

Fast, small, and fully autonomous AI personal assistant infrastructure, any OS, any platform — deploy anywhere, swap anything 🦀

Rust 32,612 4,895 Updated Aug 18, 2026

Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞

TypeScript 386,615 81,243 Updated Aug 18, 2026

MLE-bench is a benchmark for measuring how well AI agents perform at machine learning engineering

Python 1,698 257 Updated Apr 24, 2026

The official code for paper "Token-Level Collaborative Alignment for LLM-based Generative Recommendation"

Python 19 Updated Jun 9, 2026

[TMLR 26]: "UniRec: Unified Multimodal Encoding for LLM-Based Recommendations", Zijie Lei, Tao Feng, Zhigang Hua, Yan Xie, Guanyu Lin, Shuang Yang, Ge Liu, Jiaxuan You

Python 15 1 Updated Jul 9, 2026

"Unleashing the Potential of Sparse Attention on Long-term Behaviors for CTR Prediction." In Proceedings of WWW '26.

Python 9 3 Updated Jan 23, 2026

Our first fully AI generated deep learning system

Python 635 49 Updated Feb 2, 2026

Algorithm powering the For You feed on X

Rust 31,882 5,233 Updated Aug 17, 2026

High-quality single file implementation of Deep Reinforcement Learning algorithms with research-friendly features (PPO, DQN, C51, DDPG, TD3, SAC, PPG)

Python 10,280 1,154 Updated Apr 20, 2026

An Open Foundation Model and Benchmark to Accelerate Generative Recommendation

Python 899 132 Updated May 18, 2026

A compact implementation of SGLang, designed to demystify the complexities of modern LLM serving systems.

Python 4,766 789 Updated May 17, 2026

The official implementation of "ML-Master: Towards AI-for-AI via Integration of Exploration and Reasoning"

Python 441 57 Updated Mar 29, 2026
Next