← Zhiqiang He

Download PDF

Curriculum Vitae

Zhiqiang He · 何志强

Ph.D. researcher · University of Electro-Communications, Tokyo

tinyzqh@gmail.com · Google Scholar · GitHub · Zhihu

Reinforcement learning researcher working on plasticity, world models, and multi-agent RL, with publications in IEEE TMM, IEEE Communications Surveys & Tutorials, Pattern Recognition, Information Sciences, and Communications in Transportation Research. Industry experience at Baidu and InspirAI shipping RL agents into production.

Education

  • 2024 — present

    Ph.D. in Information Science

    University of Electro-Communications (UEC), Tokyo

    Advised by Prof. Zhi Liu.

  • 2019 — 2022

    M.S. in Control Science and Engineering

    Northeastern University (NEU), Shenyang

    Advised by Prof. Jiao Wang. GPA 3.29 / 4.

  • 2015 — 2019

    B.S. in Automation

    East China Jiaotong University (ECJTU), Nanchang

    Outstanding Graduate (Top 1%). GPA 3.42 / 4.

Experience

  • Jun 2022 — May 2023

    Reinforcement Learning Algorithms Engineer · InspirAI

    Hangzhou, China · Top-Performing Team Prize

    • — Built a general-purpose card-game AI SDK deployed across Sanguosha, Hearthstone, Landlord (Dou Dizhu), and GuanDan.
    • — On Landlord (Dou Dizhu), the deployed agent reached super-human level, defeating top-ranked professional players.
    • — On GuanDan, drove a +6% win-rate improvement over the previous production baseline.
  • Jun 2021 — Oct 2021

    Reinforcement Learning Research Intern · Baidu

    Beijing, China · Super Special Offer

    • — Proposed and shipped EDA-MAPPO (Expert-Data-Assisted MAPPO) into a client production environment.

Publications

  1. 01
    Plasticity-Aware Mixture of Experts for Learning Under QoE Shifts in Adaptive Video Streaming

    Zhiqiang He, Zhi Liu

    IEEE Transactions on Multimedia, 2026 · IF 9.7 · JCR Q1

  2. 02
    DiPerceiveNet: A bidirectional cross-scale perception network for vehicle re-identification

    Jihao Cai, Zhiqiang He, Zhi Liu, Yangjie Cao

    Pattern Recognition, 2026 · IF 7.6 · JCR Q1

  3. 03
    Plasticity-Enhanced Multi-Agent Mixture of Experts for Dynamic Objective Adaptation in UAV-Assisted Emergency Communication Networks

    Wen Qiu, Zhiqiang He, Wei Zhao, Hiroshi Masui

    IEEE Internet of Things Journal, 2026 · IF 8.7 · JCR Q1

  4. 04
  5. 05
    Scalable and Reliable Multi-agent Reinforcement Learning for Traffic Assignment

    Leizhen Wang, Peibo Duan, Cheng Lyu, Zewen Wang, Zhiqiang He, Nan Zheng, Zhenliang Ma

    Communications in Transportation Research, 2025 · IF 14.5 · JCR Q1

  6. 06
    A Survey on DRL based UAV Communications and Networking: DRL Fundamentals, Applications and Implementations

    Wei Zhao, Shaoxin Cui, Wen Qiu*, Zhiqiang He*, Zhi Liu, Xiao Zheng, Bomin Mao, Nei Kato

    IEEE Communications Surveys & Tutorials, 2025 · IF 42.8 · JCR Q1

  7. 07
    Understanding World Models through Multi-Step Pruning Policy via Reinforcement Learning

    Zhiqiang He, Wen Qiu, Wei Zhao, Xun Shao, Zhi Liu

    Information Sciences, 2024 · IF 8.1 · JCR Q1

  8. 08
    Erlang Planning Network: An iterative model-based reinforcement learning with multi-perspective

    Jiao Wang, Lemin Zhang, Zhiqiang He, Can Zhu, Zihui Zhao

    Pattern Recognition, 2022 · IF 8.5 · JCR Q1

  9. 09
    Control Strategy of Speed Servo Systems Based on Deep Reinforcement Learning

    Pengzhan Chen, Zhiqiang He, Chuanxi Chen, Jiahong Xu

    Algorithms, 2018

Awards

  • JST Next-Generation Researcher · ¥2.2M / year, 2025-2027
  • Outstanding Graduate (Top 1%), East China Jiaotong University, 2019
  • Honorable Mention, Mathematical Contest in Modeling (MCM), 2018
  • Third Prize, 15th Challenge Cup, Jiangxi Division, 2017

Service

Peer reviewer for

  • ACM International Conference on Multimedia (ACM MM 2026)
  • IEEE Transactions on Multimedia
  • IEEE Transactions on Network Science and Engineering
  • IEEE Internet of Things Journal
  • IEEE Open Journal of the Computer Society

Last updated 2026-07-20 · Download PDF