Skip to content
View tinyzqh's full-sized avatar

Block or report tinyzqh

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please donโ€™t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this userโ€™s behavior. Learn more about reporting abuse.

Report abuse
tinyzqh/README.md

Hi, I'm Zhiqiang He (ไฝ•ๅฟ—ๅผบ) ๐Ÿ‘‹

Building reliable agents under uncertainty.

Homepage Email Scholar Zhihu GitHub Rank Total Stars


๐ŸŽ“ About Me

  • ๐Ÿง‘โ€๐Ÿ”ฌ Ph.D. researcher at University of Electro-Communications (UEC), Tokyo (Apr 2024 โ€“ 2026)
  • ๐ŸŽฏ Working on Reinforcement Learning, with current focus on plasticity, world models, multi-agent RL, and real-system deployment
  • ๐Ÿ† JST Next-Generation Researcher (2025 โ€“ 2027)
  • ๐Ÿ’ผ Past: RL Algorithm Engineer @ InspirAI ยท RL Research Intern @ Baidu
  • โœ๏ธ I write technical notes on Zhihu โ€” 10K+ followers

๐Ÿ“Œ Selected Publications

  • Plasticity-Aware Mixture of Experts for Learning Under QoE Shifts in Adaptive Video Streaming โ€” IEEE TMM, 2026
  • A Survey on DRL based UAV Communications and Networking โ€” IEEE COMST, 2025 (co-authored)
  • Understanding World Models through Multi-Step Pruning Policy via Reinforcement Learning โ€” Information Sciences, 2024

โ†’ Full list on Google Scholar


๐Ÿ› ๏ธ Tech Stack

Python C++ C# PyTorch LaTeX Linux Git


๐Ÿ† Trophies

trophies

Pinned Loading

  1. light_mappo light_mappo Public

    Lightweight version of MAPPO to help you quickly migrate to your local environment.

    Python 875 118

  2. Systems-Intelligent-Lab/PA-MoE Systems-Intelligent-Lab/PA-MoE Public

    [IEEE TMM ACCEPTED] Official implementation of Plasticity-Aware Mixture of Experts for Learning Under QoE Shifts in Adaptive Video Streaming in IEEE Transactions on Multimedia.

    Python 3

  3. MSPP MSPP Public

    [IS] Official implementation of "Understanding world models through multi-step pruning policy via reinforcement learning" in Information Sciences.

    Python 8 3

  4. control-of-jump-systems-based-on-reinforcement-learning control-of-jump-systems-based-on-reinforcement-learning Public

    [Algorithms] Official implementation of โ€œControl Strategy of Speed Servo Systems Based on Deep Reinforcement Learningโ€

    Python 25 4

  5. awesome-reinforcement-learning awesome-reinforcement-learning Public

    Learning Resources And Links Of Reinforcement Learning ๏ผˆupdating๏ผ‰

    Python 295 83

  6. Opencv-Computer-Vision-Practice-Python- Opencv-Computer-Vision-Practice-Python- Public

    OpenCV Computer Vision Practice (Python)

    Python 194 53