Papers

2

Total Citations

14

H-Index

2

About

Qiwei He is a researcher in artificial intelligence, with a primary focus on deep reinforcement learning (DRL) and its application to complex, continuous control tasks. His major contributions center on advancing sample efficiency in environments with sparse rewards—a critical challenge in DRL. He is best known for developing "Soft Hindsight Experience Replay" (2020, 11 citations), which refines the classic Hindsight Experience Replay (HER) algorithm to enable more robust learning in robotic arm control and similar settings. Building on this, his work "Quantile Regression Hindsight Experience Replay" (2020, 3 citations) further integrates quantile regression to improve value estimation and policy stability. Though early in his career, He’s research addresses a fundamental bottleneck in DRL: making agents learn effectively from rare or delayed feedback. His methods have implications for robotics, autonomous systems, and any domain requiring efficient exploration. By tackling the brittleness of existing HER approaches, He is helping to pave the way for more reliable and practical reinforcement learning solutions.

Research Focus

Key Achievements

2
H-Index
2
Papers
14
Total Citations
7
Avg Citations/Paper
🏆 Most Cited Paper
Soft Hindsight Experience Replay
11 citations · 2020
📈 Most Prolific Year: 2020 (2 Papers)
🤝 Key Collaborators: 3
🏛 Institutions: University of Science and Technology of China

Top Papers

  1. 1
  2. 2

Key Collaborators

Contact & Links

Available for collaboration
Content generated · 15 days ago