Zhiqiang He 何志强
Ph.D. Researcher at the University of Electro-Communications, Tokyo, working on reinforcement learning. Previously an RL engineer at InspirAI and a research intern at Baidu.
Experience
-
2024 — present
Ph.D. in Information Science
University of Electro-Communications, Tokyo
Advised by Prof. Zhi Liu.
-
2022 — 2023
RL Algorithms Engineer
InspirAI, Hangzhou
Built a card-game AI SDK shipped across four production titles; the Landlord agent reached super-human level against top-ranked professional players.
-
2021
RL Research Intern
Baidu, Beijing
Proposed and shipped EDA-MAPPO into a client production environment.
-
2019 — 2022
M.S. in Control Science and Engineering
Northeastern University, Shenyang
Advised by Prof. Jiao Wang.
-
2015 — 2019
B.S. in Automation
East China Jiaotong University, Nanchang
Awards
- 2025 Selected as a JST Next-Generation Researcher, 2025–2027.
- 2019 Selected as Outstanding Graduate (Top 1%) at East China Jiaotong University.
Service
Peer reviewer for
- ACM International Conference on Multimedia (ACM MM 2026)
- IEEE Transactions on Multimedia
- IEEE Transactions on Network Science and Engineering
- IEEE Internet of Things Journal
- IEEE Open Journal of the Computer Society
Publications
-
Plasticity-Aware Mixture of Experts for Learning Under QoE Shifts in Adaptive Video Streaming
Zhiqiang He, Zhi Liu
IEEE Transactions on Multimedia, 2026 · IF 9.7 · JCR Q1
-
DiPerceiveNet: A bidirectional cross-scale perception network for vehicle re-identification
Jihao Cai, Zhiqiang He, Zhi Liu, Yangjie Cao
Pattern Recognition, 2026 · IF 7.6 · JCR Q1
-
Plasticity-Enhanced Multi-Agent Mixture of Experts for Dynamic Objective Adaptation in UAV-Assisted Emergency Communication Networks
Wen Qiu, Zhiqiang He, Wei Zhao, Hiroshi Masui
IEEE Internet of Things Journal, 2026 · IF 8.7 · JCR Q1
-
Silent Neuron Theory and Plasticity Preservation for Deep Reinforcement Learning in Adaptive Video Streaming
Zhiqiang He, Zhi Liu
arXiv preprint, 2025
-
Scalable and Reliable Multi-agent Reinforcement Learning for Traffic Assignment
Leizhen Wang, Peibo Duan, Cheng Lyu, Zewen Wang, Zhiqiang He, Nan Zheng, Zhenliang Ma
Communications in Transportation Research, 2025 · IF 14.5 · JCR Q1
-
A Survey on DRL based UAV Communications and Networking: DRL Fundamentals, Applications and Implementations
Wei Zhao, Shaoxin Cui, Wen Qiu*, Zhiqiang He*, Zhi Liu, Xiao Zheng, Bomin Mao, Nei Kato
IEEE Communications Surveys & Tutorials, 2025 · IF 42.8 · JCR Q1
-
Understanding World Models through Multi-Step Pruning Policy via Reinforcement Learning
Zhiqiang He, Wen Qiu, Wei Zhao, Xun Shao, Zhi Liu
Information Sciences, 2024 · IF 8.1 · JCR Q1
-
Erlang Planning Network: An iterative model-based reinforcement learning with multi-perspective
Jiao Wang, Lemin Zhang, Zhiqiang He, Can Zhu, Zihui Zhao
Pattern Recognition, 2022 · IF 8.5 · JCR Q1
-
Control Strategy of Speed Servo Systems Based on Deep Reinforcement Learning
Pengzhan Chen, Zhiqiang He, Chuanxi Chen, Jiahong Xu
Algorithms, 2018