Skip to content
View yyyouy's full-sized avatar
🏠
Working from home
🏠
Working from home
  • renmin university of china
  • beijing
  • 15:55 (UTC +08:00)

Highlights

  • Pro

Organizations

@ML-GSAI

Block or report yyyouy

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

A Next-Generation Training Engine Built for Ultra-Large MoE Models

Python 5,164 432 Updated Jul 23, 2026

Structured deep research skill for Claude Code/Open Code/Codex with human-in-the-loop control

Python 1,751 140 Updated May 7, 2026

UniRL is a Framework for Unified Multimodal Model Reinforcement Learning

Python 853 55 Updated Jul 24, 2026

Local AI filmmaking studio — skills, canvas, timeline — driven from your coding agent.

JavaScript 318 25 Updated Jul 23, 2026

Official PyTorch implementation for "Principled RL for Diffusion LLMs Emerges from a Sequence-Level Perspective"

Python 39 2 Updated Jan 25, 2026

Awesome List for On-Policy Distillation

763 18 Updated Jul 23, 2026

SenseNova-U series: Native Unified Paradigm with NEO-unify from the First Principles

Python 4,366 372 Updated Jul 16, 2026

21 writing rules for AI coding and writing agents. Drop-in for Claude Code, Codex, Copilot, Cursor, and Aider, so their output reads like a tech pro.

Python 584 32 Updated Jun 13, 2026

A Curated List of Awesome Video World Models with AR Diffusion: Covering Algorithms, Applications, and Infrastructure, Aimed at Serving as a Comprehensive Resource for Researchers, Practitioners, a…

TeX 673 18 Updated Jun 4, 2026

One config to rule all your AI agents: portable (every project, every session), effective (curated writing, routing, skills), and safer (destructive-command guard).

Python 199 23 Updated Jul 24, 2026

Agent skill for harness engineering — memory, permissions, context engineering, multi-agent coordination. Distilled from Claude Code, with Codex CLI and Gemini CLI on the roadmap. EN/ZH. Install vi…

296 48 Updated Apr 2, 2026

slime is an LLM post-training framework for RL Scaling.

Python 7,625 1,095 Updated Jul 24, 2026

[NeurIPS 2025] Beyond Masked and Unmasked: Discrete Diffusion Models via Partial Masking

Python 32 1 Updated Jun 15, 2026

Democratizing Reinforcement Learning for LLMs

Python 5,730 594 Updated Jul 25, 2026

Build your own Claude Code from scratch. 🔍 Claude Code 开源了 50 万行代码,读不动?用 ~5000 行 TypeScript / Python 从零复现核心架构,11 章分步教程带你理解 coding agent 精髓

Python 2,461 509 Updated Jul 9, 2026

Deep dive into Claude Code internals — architecture, agent loop, context engineering, and more. / 深入解析 Claude Code 源码:架构、Agent 循环、上下文工程、工具系统等

3,285 682 Updated Jul 12, 2026

OmX - Oh My codeX: Your codex is not alone. Add hooks, agent teams, HUDs, and so much more.

TypeScript 32,221 2,498 Updated Jul 25, 2026

An agent-managed museum exhibit, built in Rust with Gajae-Code / LazyCodex — developed and maintained with no human intervention.

Rust 194,891 109,501 Updated Jun 26, 2026

Kimi Code CLI is your next CLI agent.

Python 10,805 1,251 Updated Jul 16, 2026

📚 《从零开始构建智能体》——从零开始的智能体原理与实践教程

Python 68,500 8,526 Updated Jul 17, 2026

给 Claude Code 装上完整联网能力的 skill:三层通道调度 + 浏览器 CDP + 并行分治

JavaScript 8,424 596 Updated May 16, 2026

Public repository for Agent Skills

Python 164,019 19,472 Updated Jul 24, 2026

Open source repository of plugins primarily intended for knowledge workers to use in Claude Cowork

Python 23,026 2,762 Updated Jul 25, 2026

verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework

Python 22,652 4,274 Updated Jul 25, 2026

Official repository of DARE: Diffusion Large Language Models Alignment and Reinforcement Executor

Python 214 7 Updated Jul 18, 2026

ARIS ⚔️ (Auto-Research-In-Sleep) — Lightweight Markdown-only skills for autonomous ML research: cross-model review loops, idea discovery, and experiment automation. No framework, no lock-in — works…

Python 13,829 1,238 Updated Jul 22, 2026
Python 53 4 Updated May 16, 2026

A unified framework for easy reinforcement learning in Flow-Matching models

Python 640 51 Updated Jul 12, 2026
Next