Zhiqiang He

Zhiqiang He 何志强

Ph.D. Researcher at the University of Electro-Communications, Tokyo, working on reinforcement learning. Previously an RL engineer at InspirAI and a research intern at Baidu.

Experience

  1. 2024 — present

    Ph.D. in Information Science

    University of Electro-Communications, Tokyo

    Advised by Prof. Zhi Liu.

  2. 2022 — 2023

    RL Algorithms Engineer

    InspirAI, Hangzhou

    Built a card-game AI SDK shipped across four production titles; the Landlord agent reached super-human level against top-ranked professional players.

  3. 2021

    RL Research Intern

    Baidu, Beijing

    Proposed and shipped EDA-MAPPO into a client production environment.

  4. 2019 — 2022

    M.S. in Control Science and Engineering

    Northeastern University, Shenyang

    Advised by Prof. Jiao Wang.

  5. 2015 — 2019

    B.S. in Automation

    East China Jiaotong University, Nanchang

Awards

  1. 2025 Selected as a JST Next-Generation Researcher, 2025–2027.
  2. 2019 Selected as Outstanding Graduate (Top 1%) at East China Jiaotong University.

Service

Peer reviewer for

  • ACM International Conference on Multimedia (ACM MM 2026)
  • IEEE Transactions on Multimedia
  • IEEE Transactions on Network Science and Engineering
  • IEEE Internet of Things Journal
  • IEEE Open Journal of the Computer Society

Publications

  1. Plasticity-Aware Mixture of Experts for Learning Under QoE Shifts in Adaptive Video Streaming

    Zhiqiang He, Zhi Liu

    IEEE Transactions on Multimedia, 2026 · IF 9.7 · JCR Q1

    Paper arXiv Code

  2. DiPerceiveNet: A bidirectional cross-scale perception network for vehicle re-identification

    Jihao Cai, Zhiqiang He, Zhi Liu, Yangjie Cao

    Pattern Recognition, 2026 · IF 7.6 · JCR Q1

    Paper PDF

  3. Plasticity-Enhanced Multi-Agent Mixture of Experts for Dynamic Objective Adaptation in UAV-Assisted Emergency Communication Networks

    Wen Qiu, Zhiqiang He, Wei Zhao, Hiroshi Masui

    IEEE Internet of Things Journal, 2026 · IF 8.7 · JCR Q1

    Paper arXiv

  4. Silent Neuron Theory and Plasticity Preservation for Deep Reinforcement Learning in Adaptive Video Streaming

    Zhiqiang He, Zhi Liu

    arXiv preprint, 2025

    Paper arXiv

  5. Scalable and Reliable Multi-agent Reinforcement Learning for Traffic Assignment

    Leizhen Wang, Peibo Duan, Cheng Lyu, Zewen Wang, Zhiqiang He, Nan Zheng, Zhenliang Ma

    Communications in Transportation Research, 2025 · IF 14.5 · JCR Q1

    Paper PDF Code

  6. A Survey on DRL based UAV Communications and Networking: DRL Fundamentals, Applications and Implementations

    Wei Zhao, Shaoxin Cui, Wen Qiu*, Zhiqiang He*, Zhi Liu, Xiao Zheng, Bomin Mao, Nei Kato

    IEEE Communications Surveys & Tutorials, 2025 · IF 42.8 · JCR Q1

    Paper

  7. Understanding World Models through Multi-Step Pruning Policy via Reinforcement Learning

    Zhiqiang He, Wen Qiu, Wei Zhao, Xun Shao, Zhi Liu

    Information Sciences, 2024 · IF 8.1 · JCR Q1

    Paper PDF Code

  8. Erlang Planning Network: An iterative model-based reinforcement learning with multi-perspective

    Jiao Wang, Lemin Zhang, Zhiqiang He, Can Zhu, Zihui Zhao

    Pattern Recognition, 2022 · IF 8.5 · JCR Q1

    Paper PDF

  9. Control Strategy of Speed Servo Systems Based on Deep Reinforcement Learning

    Pengzhan Chen, Zhiqiang He, Chuanxi Chen, Jiahong Xu

    Algorithms, 2018

    Paper Code