-
The Chinese University of Hong Kong, Shenzhen
- Shenzhen
- www.liziniu.org
- @ziniuli
Highlights
- Pro
Stars
Model Context Protocol (MCP) server that lets AI assistants read Overleaf projects, parse LaTeX document structure, and push section-level edits back via Git. Compatible with Claude Desktop, Cursor…
Long-horizon agent control plane for durable, governed work across Codex, Claude Code, and other harnesses.
Symphony turns project work into isolated, autonomous implementation runs, allowing teams to manage work instead of supervising coding agents.
A next.js web application that integrates AI capabilities with draw.io diagrams. This app allows you to create, modify, and enhance diagrams through natural language commands and AI-assisted visual…
Get the main content of any page as Markdown.
Uni-Agent is a framework for training long-horizon agents.
Unlimited OCR Works: Welcome the Era of One-shot Long-horizon Parsing.
code for "Tying the Loop - Tied Expert Layers in Mixture-of-Experts Language Models"
Code and Data for paper "GameCraft-Bench: Can Agents Build Playable Games End-to-End in a Real Game Engine?"
Lightweight coding agent that runs in your terminal
Skill to give Claude Code (and any coding agent) the ability to generate beautiful and practical Excalidraw diagrams.
DeepGEMM: clean and efficient BLAS kernel library on GPU
verl Zero-Mismatch Dense/MoE HuggingFace Rollout
A unified library of SOTA model optimization techniques like quantization, distillation, pruning, neural architecture search, speculative decoding, etc. It compresses deep learning models for downs…
A SOTA quantization algorithm for high-accuracy low-bit LLM inference, seamlessly optimized for CPU/XPU/CUDA, with multi-datatype support and full compatibility with vLLM, SGLang, and Transformers.
An asynchronous streaming data management module for efficient post-training.
Experimenting Heuristic Learning with ImageNet
MyPhoneBench: Do Phone-Use Agents Respect Your Privacy?
A MemAgent framework that can be extrapolated to 3.5M, along with a training framework for RL training of any agent workflow.
[ICLR'26] Stronger-MAS: A RL Framework for multi LLM agent system; [arxiv] MetaAgent-X: End-to-End Reinforcement Learning Automatic Multi-Agent Systems
Checkpoint-engine is a simple middleware to update model weights in LLM inference engines
Code and implementations for the paper "AgentGym-RL: Training LLM Agents for Long-Horizon Decision Making through Multi-Turn Reinforcement Learning" by Zhiheng Xi et al.
Code for ICLR 2026 paper on "Understanding and Improving Shampoo and SOAP via Kullback-Leibler Minimization"
Implementation of GradLoc from the Tencent Hunyuan blog "Stabilizing RLVR via Token-level Gradient Diagnosis and Layerwise Clipping".