Skip to content
View duoan's full-sized avatar
🎯
Focusing
🎯
Focusing

Block or report duoan

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Megatron's multi-modal data loader

Python 377 56 Updated Aug 4, 2026

Code and Data for Tau-Bench

Python 1,366 213 Updated Mar 18, 2026

Agentic RL最详细入门

HTML 261 13 Updated Aug 6, 2026

High-performance RL post-training infrastructure. Designed to achieve bitwise operator-level train-inference consistency across heterogeneous engines and extreme memory efficiency for GRPO, PPO, etc.

Python 222 62 Updated Aug 8, 2026

SWE-agent takes a GitHub issue and tries to automatically fix it, using your LM of choice. It can also be employed for offensive cybersecurity or competitive coding challenges. [NeurIPS 2024]

Python 20,025 2,185 Updated Aug 3, 2026

🙌 OpenHands: AI-Driven Development

TypeScript 83,484 10,789 Updated Aug 8, 2026

Secure and fast microVMs for serverless computing.

Rust 35,953 2,546 Updated Aug 7, 2026

A curated list of reinforcement learning with verifiable rewards (continually updated)

290 21 Updated Jun 1, 2026

DLRover: An Automatic Distributed Deep Learning System

Python 1,672 215 Updated Jul 29, 2026

The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.

Python 5,648 573 Updated Aug 7, 2026

Our library for RL environments + evals

Python 4,475 635 Updated Aug 8, 2026

An asynchronous streaming data management module for efficient post-training.

Python 131 43 Updated Aug 8, 2026

RLinf: Reinforcement Learning Infrastructure for Embodied and Agentic AI

Python 4,481 641 Updated Aug 8, 2026

Model interpretability and understanding for PyTorch

Python 5,683 560 Updated Jul 31, 2026

The WeightWatcher tool for predicting the accuracy of Deep Neural Networks

Python 1,771 146 Updated May 11, 2026

A flexible and high-performance training framework designed for large-scale foundation model training on AMD GPUs

Python 119 48 Updated Aug 8, 2026

High-performance GPU kernels for LLM inference in OpenAI Triton. Fused RMSNorm, SwiGLU, INT8 GEMM with benchmarks and roofline analysis.

Python 38 6 Updated Jul 22, 2026

Code release for book "Efficient Training in PyTorch"

Python 133 21 Updated Apr 10, 2025

AI 基础知识 - GPU 架构、CUDA 编程、大模型基础及AI Agent 相关知识。

HTML 2,149 331 Updated Aug 7, 2026

TradingAgents: Multi-Agents LLM Financial Trading Framework

Python 96,390 18,642 Updated Jul 18, 2026

AI Infrastructure Engineer Learning Track - Production ML infrastructure curriculum (2-4 years experience)

Python 1,559 262 Updated Jun 26, 2026

Academic Research Skills for Claude Code: research → write → review → revise → finalize

Python 41,391 3,295 Updated Aug 8, 2026

Now, Stronger AI Pushes Frontiers, Stronger Our Shared Future.

TypeScript 3,256 331 Updated Jun 28, 2026

ARIS ⚔️ (Auto-Research-In-Sleep) — Lightweight Markdown-only skills for autonomous ML research: cross-model review loops, idea discovery, and experiment automation. No framework, no lock-in — works…

Python 14,411 1,281 Updated Aug 7, 2026

A curriculum for learning about gpu performance engineering, from scratch to what the frontier AI labs do

1,288 162 Updated Apr 27, 2026

AI Infrastructure Performance Engineer Learning Track - GPU optimization, inference optimization, and cost reduction

Python 53 11 Updated Jul 8, 2026

Universal LLM Deployment Engine with ML Compilation

Python 23,046 2,110 Updated Jul 31, 2026

An Extensible Deep Learning Library

Python 2,371 409 Updated Jul 8, 2026

Helpful kernel tutorials, examples and SKILLs for tile-based GPU programming

Python 795 82 Updated Aug 8, 2026
Next