-
Tsinghua University
- Shenzhen, China
-
11:17
(UTC +08:00) - https://www.tonyzhao.xyz
Stars
A user-friendly & efficient knowledge distillation framework for LLMs, supporting off-policy, on-policy (OPD), cross-tokenizer, multimodal, and on-policy self-distillation.
[ICLR 2026 Blogpost Track Poster] JustRL: Scaling a 1.5B LLM with a Simple RL Recipe
Task-oriented AI Agent productivity platform
The code of EfficientPosterGen: Semantic-aware Efficient Poster Generation via Token Compression and Accurate Violation Detection
A lightweight, AI-native training framework for large language models. Designed for fast iteration, reproducible experiments, and modular configuration across SFT, RLVR, and evaluation workflows.
lijinnn / MagiAttention
Forked from SandAI-org/MagiAttentionA Distributed Attention Towards Linear Scalability for Ultra-Long Context, Heterogeneous Data Training
A curated list of awesome resources about reward construction for AI agents. This repository covers cutting-edge research, and practical guides on defining and collecting rewards to build more inte…
🚀 Awesome System for Machine Learning ⚡️ AI System Papers and Industry Practice. ⚡️ System for Machine Learning, LLM (Large Language Model), GenAI (Generative AI). 🍻 OSDI, NSDI, SIGCOMM, SoCC, MLSy…
A fancy self-hosted monitoring tool
Trae Agent is an LLM-based agent for general purpose software engineering tasks.
SGLang is a high-performance serving framework for large language models and multimodal models.
tufu9441 / maupassant-hexo
Forked from icylogic/maupassant-hexoA simple Hexo theme forked from icylogic.
Large Language Model (LLM) Systems Paper List
Tensors and Dynamic neural networks in Python with strong GPU acceleration
VeOmni: Scaling Any Modality Model Training with Model-Centric Distributed Recipe Zoo
Curated collection of papers in machine learning systems
verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework
My learning notes for ML SYS.