-
The Hong Kong University of Science and Technology
- HongKong
-
01:46
(UTC +08:00) - https://zijianzhao.netlify.app/
- https://orcid.org/0000-0002-3326-9650
- https://scholar.google.com/citations?user=XkA3qCcAAAAJ&hl=en
- https://openreview.net/profile?id=~Zijian_Zhao7
- https://huggingface.co/RS2002
Stars
Skill package for ML/CV/NLP paper writing, curated and adapted from Prof. Peng Sida's open notes for Codex, Claude Code, and Gemini.
Code for tasks on Cainiao-LaDe (Last-mile Delivery dataset).
Implementation for “Hierarchical Optimization via LLM-Guided Objective Evolution for Mobility-on-Demand Systems.” (NeurIPS 2025)
xingtian is a componentized library for the development and verification of reinforcement learning algorithms
A next-generation LLM4AD platform focused on intuitive UI interactions and seamless collaboration with AI agents, making automated algorithm design more accessible and easier to use
Code for paper "SPG Sandwiched Policy Gradient for Masked Diffusion Language Models"
Official implementation of "Fast-dLLM: Training-free Acceleration of Diffusion LLM by Enabling KV Cache and Parallel Decoding"
To make music production easier, we introduce Amadeus , a novel MIDI generation framework. While significantly improving generation quality, we have achieved a speedup of at least 4x compared to pu…
Official implementation of "Diffusion Language Models Know the Answer Before Decoding"
A PyTorch library for all things Reinforcement Learning (RL) for Combinatorial Optimization (CO)
[EMNLP 2024 (main)] Attention Score is not All You Need for Token Importance Indicator in KV Cache Reduction: Value Also Matters
BertViz: Visualize Attention in Transformer Models
[ICML2026] Official Pytorch Implement for "Search or Accelerate: Confidence-Switched Position Beam Search for Diffusion Language Models"
Implementations of IQL, QMIX, VDN, COMA, QTRAN, MAVEN, CommNet, DyMA-CL, and G2ANet on SMAC, the decentralised micromanagement scenario of StarCraft II
Train transformer language models with reinforcement learning.
multi-agent deep reinforcement learning for networked system control.
Official implementation for "Unifying Masked Diffusion Models with Various Generation Orders and Beyond"
A framework for few-shot evaluation of language models.
Official Implementation for the paper "d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning"
Revisiting Discrete Gradient Estimation in MADDPG
Official PyTorch implementation for ICLR2025 paper "Scaling up Masked Diffusion Models on Text"
Multi-Agent Reinforcement Learning (MARL) papers