Curriculum Vitae
Zhiqiang He · 何志强
Ph.D. researcher · University of Electro-Communications, Tokyo
tinyzqh@gmail.com · Google Scholar · GitHub · Zhihu
Reinforcement learning researcher working on plasticity, world models, and multi-agent RL, with publications in IEEE TMM, IEEE Communications Surveys & Tutorials, Pattern Recognition, Information Sciences, and Communications in Transportation Research. Industry experience at Baidu and InspirAI shipping RL agents into production.
Education
- 2024 — present
Ph.D. in Information Science
University of Electro-Communications (UEC), Tokyo
Advised by Prof. Zhi Liu.
- 2019 — 2022
M.S. in Control Science and Engineering
Northeastern University (NEU), Shenyang
Advised by Prof. Jiao Wang. GPA 3.29 / 4.
- 2015 — 2019
B.S. in Automation
East China Jiaotong University (ECJTU), Nanchang
Outstanding Graduate (Top 1%). GPA 3.42 / 4.
Experience
- Jun 2022 — May 2023
Reinforcement Learning Algorithms Engineer · InspirAI
Hangzhou, China · Top-Performing Team Prize
- — Built a general-purpose card-game AI SDK deployed across Sanguosha, Hearthstone, Landlord (Dou Dizhu), and GuanDan.
- — On Landlord (Dou Dizhu), the deployed agent reached super-human level, defeating top-ranked professional players.
- — On GuanDan, drove a +6% win-rate improvement over the previous production baseline.
- Jun 2021 — Oct 2021
Reinforcement Learning Research Intern · Baidu
Beijing, China · Super Special Offer
- — Proposed and shipped EDA-MAPPO (Expert-Data-Assisted MAPPO) into a client production environment.
Publications
- 01 Plasticity-Aware Mixture of Experts for Learning Under QoE Shifts in Adaptive Video Streaming
Zhiqiang He, Zhi Liu
IEEE Transactions on Multimedia, 2026 · IF 9.7 · JCR Q1
- 02 DiPerceiveNet: A bidirectional cross-scale perception network for vehicle re-identification
Jihao Cai, Zhiqiang He, Zhi Liu, Yangjie Cao
Pattern Recognition, 2026 · IF 7.6 · JCR Q1
- 03 Plasticity-Enhanced Multi-Agent Mixture of Experts for Dynamic Objective Adaptation in UAV-Assisted Emergency Communication Networks
Wen Qiu, Zhiqiang He, Wei Zhao, Hiroshi Masui
IEEE Internet of Things Journal, 2026 · IF 8.7 · JCR Q1
- 04 Silent Neuron Theory and Plasticity Preservation for Deep Reinforcement Learning in Adaptive Video Streaming
Zhiqiang He, Zhi Liu
arXiv preprint, 2025
- 05 Scalable and Reliable Multi-agent Reinforcement Learning for Traffic Assignment
Leizhen Wang, Peibo Duan, Cheng Lyu, Zewen Wang, Zhiqiang He, Nan Zheng, Zhenliang Ma
Communications in Transportation Research, 2025 · IF 14.5 · JCR Q1
- 06 A Survey on DRL based UAV Communications and Networking: DRL Fundamentals, Applications and Implementations
Wei Zhao, Shaoxin Cui, Wen Qiu*, Zhiqiang He*, Zhi Liu, Xiao Zheng, Bomin Mao, Nei Kato
IEEE Communications Surveys & Tutorials, 2025 · IF 42.8 · JCR Q1
- 07 Understanding World Models through Multi-Step Pruning Policy via Reinforcement Learning
Zhiqiang He, Wen Qiu, Wei Zhao, Xun Shao, Zhi Liu
Information Sciences, 2024 · IF 8.1 · JCR Q1
- 08 Erlang Planning Network: An iterative model-based reinforcement learning with multi-perspective
Jiao Wang, Lemin Zhang, Zhiqiang He, Can Zhu, Zihui Zhao
Pattern Recognition, 2022 · IF 8.5 · JCR Q1
- 09 Control Strategy of Speed Servo Systems Based on Deep Reinforcement Learning
Pengzhan Chen, Zhiqiang He, Chuanxi Chen, Jiahong Xu
Algorithms, 2018
Awards
- JST Next-Generation Researcher · ¥2.2M / year, 2025-2027
- Outstanding Graduate (Top 1%), East China Jiaotong University, 2019
- Honorable Mention, Mathematical Contest in Modeling (MCM), 2018
- Third Prize, 15th Challenge Cup, Jiangxi Division, 2017
Service
Peer reviewer for
- ACM International Conference on Multimedia (ACM MM 2026)
- IEEE Transactions on Multimedia
- IEEE Transactions on Network Science and Engineering
- IEEE Internet of Things Journal
- IEEE Open Journal of the Computer Society
Last updated 2026-07-20 · Download PDF