Skip to content
View duoan's full-sized avatar
🎯
Focusing
🎯
Focusing

Block or report duoan

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

A Next-Generation Training Engine Built for Ultra-Large MoE Models

Python 5,179 444 Updated Aug 14, 2026

Megatron's multi-modal data loader

Python 377 56 Updated Aug 4, 2026

Code and Data for Tau-Bench

Python 1,380 213 Updated Mar 18, 2026

Agentic RL最详细入门

HTML 275 14 Updated Aug 6, 2026

High-performance RL post-training infrastructure. Designed to achieve bitwise operator-level train-inference consistency across heterogeneous engines and extreme memory efficiency for GRPO, PPO, etc.

Python 264 69 Updated Aug 13, 2026

SWE-agent takes a GitHub issue and tries to automatically fix it, using your LM of choice. It can also be employed for offensive cybersecurity or competitive coding challenges. [NeurIPS 2024]

Python 20,055 2,194 Updated Aug 10, 2026

🙌 OpenHands: AI-Driven Development

TypeScript 84,007 10,885 Updated Aug 14, 2026

Secure and fast microVMs for serverless computing.

Rust 36,049 2,557 Updated Aug 14, 2026

A curated list of reinforcement learning with verifiable rewards (continually updated)

295 22 Updated Jun 1, 2026

DLRover: An Automatic Distributed Deep Learning System

Python 1,674 216 Updated Aug 14, 2026

The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.

Python 5,667 574 Updated Aug 14, 2026

Our library for RL environments + evals

Python 4,509 643 Updated Aug 14, 2026

An asynchronous streaming data management module for efficient post-training.

Python 133 44 Updated Aug 12, 2026

RLinf: Reinforcement Learning Infrastructure for Embodied and Agentic AI

Python 4,533 654 Updated Aug 13, 2026

Model interpretability and understanding for PyTorch

Python 5,686 563 Updated Aug 13, 2026

The WeightWatcher tool for predicting the accuracy of Deep Neural Networks

Python 1,771 146 Updated May 11, 2026

A flexible and high-performance training framework designed for large-scale foundation model training on AMD GPUs

Python 119 46 Updated Aug 14, 2026

High-performance GPU kernels for LLM inference in OpenAI Triton. Fused RMSNorm, SwiGLU, INT8 GEMM with benchmarks and roofline analysis.

Python 38 7 Updated Jul 22, 2026

Code release for book "Efficient Training in PyTorch"

Python 133 21 Updated Apr 10, 2025

AI 基础知识 - GPU 架构、CUDA 编程、大模型基础及AI Agent 相关知识。

HTML 2,237 348 Updated Aug 11, 2026

TradingAgents: Multi-Agents LLM Financial Trading Framework

Python 98,087 18,885 Updated Jul 18, 2026

AI Infrastructure Engineer Learning Track - Production ML infrastructure curriculum (2-4 years experience)

Python 1,597 273 Updated Jun 26, 2026

Academic Research Skills for Claude Code: research → write → review → revise → finalize

Python 42,454 3,381 Updated Aug 14, 2026

Now, Stronger AI Pushes Frontiers, Stronger Our Shared Future.

TypeScript 3,269 331 Updated Jun 28, 2026

ARIS ⚔️ (Auto-Research-In-Sleep) — Lightweight Markdown-only skills for autonomous ML research: cross-model review loops, idea discovery, and experiment automation. No framework, no lock-in — works…

Python 14,671 1,293 Updated Aug 11, 2026

A curriculum for learning about gpu performance engineering, from scratch to what the frontier AI labs do

1,310 164 Updated Apr 27, 2026

AI Infrastructure Performance Engineer Learning Track - GPU optimization, inference optimization, and cost reduction

Python 55 12 Updated Jul 8, 2026

Universal LLM Deployment Engine with ML Compilation

Python 23,064 2,112 Updated Jul 31, 2026

An Extensible Deep Learning Library

Python 2,373 409 Updated Jul 8, 2026
Next