Skip to content
View JunnYu's full-sized avatar
🐢
Focusing
🐢
Focusing

Organizations

@PaddlePaddle

Block or report JunnYu

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

MAGI-2-preview: Scaling Video Generation Models Efficiently

Python 519 12 Updated Aug 6, 2026
Python 910 87 Updated Aug 15, 2026

Uni-Agent is a framework for training long-horizon agents.

Python 510 93 Updated Aug 14, 2026

Miles is an enterprise-facing reinforcement learning framework for LLM and VLM post-training, forked from and co-evolving with slime.

Python 1,999 360 Updated Aug 15, 2026

Build compute kernels and load them from the Hub.

Python 722 120 Updated Aug 15, 2026

The agent harness performance optimization system. Skills, instincts, memory, security, and research-first development for Claude Code, Codex, Opencode, Cursor and beyond.

JavaScript 240,274 36,463 Updated Aug 15, 2026

An asynchronous streaming data management module for efficient post-training.

Python 133 44 Updated Aug 12, 2026

A high-performance RL training-inference weight synchronization framework, designed to enable second-level parameter updates from training to inference in RL workflows

Python 169 21 Updated Aug 14, 2026

Contexts Optical Compression

Python 23,800 2,197 Updated Jan 27, 2026

Agent Reinforcement Trainer: train multi-step agents for real-world tasks using GRPO. Give your agents on-the-job training. Reinforcement learning for Qwen3.6, GPT-OSS, Llama, and more!

Python 10,582 975 Updated Aug 15, 2026

An Efficient and User-Friendly Scaling Library for Reinforcement Learning with Large Language Models

Python 3,362 305 Updated Aug 15, 2026

Agentic RL Training at Scale

Python 1,925 404 Updated Aug 15, 2026

A PyTorch native platform for training generative AI models

Python 5,624 953 Updated Aug 15, 2026

诺亚盘古大模型研发背后的真正的心酸与黑暗的故事。

11,541 1,300 Updated Jul 9, 2025

Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.

Python 87,698 11,176 Updated Jul 22, 2026

The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.

Python 5,667 574 Updated Aug 15, 2026

Unleashing the Power of Reinforcement Learning for Math and Code Reasoners

Python 738 44 Updated Jun 6, 2025

Distributed Compiler based on Triton for Parallel Systems

Python 1,518 165 Updated Aug 12, 2026

FlashInfer: Kernel Library for LLM Serving

Python 6,169 1,280 Updated Aug 15, 2026

TensorDict is a pytorch dedicated tensor container.

Python 1,034 116 Updated Aug 15, 2026

A high-performance distributed file system designed to address the challenges of AI training and inference workloads.

C++ 10,124 1,080 Updated May 7, 2026

A bidirectional pipeline parallelism algorithm for computation-communication overlap in DeepSeek V3/R1 training.

Python 2,991 327 Updated Jan 14, 2026

DeepGEMM: clean and efficient BLAS kernel library on GPU

Cuda 7,683 1,172 Updated Aug 11, 2026

DeepEP: an efficient expert-parallel communication library

Cuda 9,992 1,380 Updated Aug 5, 2026

FlashMLA: Efficient Multi-head Latent Attention Kernels

C++ 12,845 1,123 Updated Jul 28, 2026

Official Repo for Open-Reasoner-Zero

Python 2,098 120 Updated Jun 2, 2025

A stand-alone implementation of several NumPy dtype extensions used in machine learning.

C++ 357 59 Updated Aug 14, 2026

Solve Visual Understanding with Reinforced VLMs

Python 6,016 385 Updated Jul 7, 2026

Integrate the DeepSeek API into popular software

38,761 4,221 Updated Feb 23, 2026

User-friendly AI Interface (Supports Ollama, OpenAI API, ...)

Python 148,862 21,676 Updated Aug 15, 2026
Next