Skip to content
View JunnYu's full-sized avatar
🐢
Focusing
🐢
Focusing

Organizations

@PaddlePaddle

Block or report JunnYu

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

MAGI-2-preview: Scaling Video Generation Models Efficiently

Python 458 12 Updated Aug 6, 2026
Python 904 86 Updated Aug 12, 2026

Uni-Agent is a framework for training long-horizon agents.

Python 507 91 Updated Aug 12, 2026

Miles is an enterprise-facing reinforcement learning framework for LLM and VLM post-training, forked from and co-evolving with slime.

Python 1,974 354 Updated Aug 12, 2026

Build compute kernels and load them from the Hub.

Python 723 119 Updated Aug 12, 2026

The agent harness performance optimization system. Skills, instincts, memory, security, and research-first development for Claude Code, Codex, Opencode, Cursor and beyond.

JavaScript 239,748 36,387 Updated Aug 12, 2026

An asynchronous streaming data management module for efficient post-training.

Python 133 44 Updated Aug 12, 2026

A high-performance RL training-inference weight synchronization framework, designed to enable second-level parameter updates from training to inference in RL workflows

Python 169 21 Updated Aug 7, 2026

Contexts Optical Compression

Python 23,761 2,194 Updated Jan 27, 2026

Agent Reinforcement Trainer: train multi-step agents for real-world tasks using GRPO. Give your agents on-the-job training. Reinforcement learning for Qwen3.6, GPT-OSS, Llama, and more!

Python 10,580 972 Updated Aug 12, 2026

An Efficient and User-Friendly Scaling Library for Reinforcement Learning with Large Language Models

Python 3,360 305 Updated Aug 12, 2026

Agentic RL Training at Scale

Python 1,898 393 Updated Aug 12, 2026

A PyTorch native platform for training generative AI models

Python 5,620 946 Updated Aug 12, 2026

诺亚盘古大模型研发背后的真正的心酸与黑暗的故事。

11,540 1,300 Updated Jul 9, 2025

Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.

Python 87,527 11,171 Updated Jul 22, 2026

The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.

Python 5,658 574 Updated Aug 12, 2026

Unleashing the Power of Reinforcement Learning for Math and Code Reasoners

Python 740 44 Updated Jun 6, 2025

Distributed Compiler based on Triton for Parallel Systems

Python 1,517 165 Updated Aug 12, 2026

FlashInfer: Kernel Library for LLM Serving

Python 6,154 1,271 Updated Aug 12, 2026

TensorDict is a pytorch dedicated tensor container.

Python 1,034 116 Updated Aug 12, 2026

A high-performance distributed file system designed to address the challenges of AI training and inference workloads.

C++ 10,110 1,076 Updated May 7, 2026

A bidirectional pipeline parallelism algorithm for computation-communication overlap in DeepSeek V3/R1 training.

Python 2,989 327 Updated Jan 14, 2026

DeepGEMM: clean and efficient BLAS kernel library on GPU

Cuda 7,656 1,164 Updated Aug 11, 2026

DeepEP: an efficient expert-parallel communication library

Cuda 9,975 1,373 Updated Aug 5, 2026

FlashMLA: Efficient Multi-head Latent Attention Kernels

C++ 12,837 1,120 Updated Jul 28, 2026

Official Repo for Open-Reasoner-Zero

Python 2,098 121 Updated Jun 2, 2025

A stand-alone implementation of several NumPy dtype extensions used in machine learning.

C++ 357 59 Updated Jun 30, 2026

Solve Visual Understanding with Reinforced VLMs

Python 6,016 385 Updated Jul 7, 2026

Integrate the DeepSeek API into popular software

38,590 4,204 Updated Feb 23, 2026

User-friendly AI Interface (Supports Ollama, OpenAI API, ...)

Python 148,612 21,632 Updated Aug 12, 2026
Next