About

Chenjia Bai is a prominent reinforcement learning researcher whose work spans exploration strategies, multi-goal learning, offline RL, and robot locomotion. He is perhaps best known for his comprehensive survey on exploration in deep reinforcement learning — covering both single-agent and multiagent domains — which has garnered an impressive 158 citations and stands as an authoritative reference for researchers tackling sample inefficiency in complex environments like game AI, autonomous vehicles, and robotics. His contributions to multi-goal reinforcement learning are equally notable, with a series of works addressing hindsight experience replay, hindsight bias, and attentive goal generation that collectively advance how agents learn from sparse-reward settings. Bai has also made meaningful strides in safety-aware offline RL, proposing monotonic quantile networks to optimize worst-case return criteria — a critical concern in safety-sensitive applications. More recently, his research has extended into physical robot control, with notable work on risk-averse quadrupedal locomotion and humanoid balance on challenging terrain, demonstrating a clear trajectory toward real-world deployment. Across his body of work, Bai consistently bridges theoretical rigor with practical impact, making him a valuable voice in modern reinforcement learning research.

Research Focus

Key Achievements

7
H-Index
11
Papers
275
Total Citations
25
Avg Citations/Paper
🏆 Most Cited Paper
Exploration in Deep Reinforcement Learning: From Single-Agent to Multiagent Domain
158 citations · 2023
📈 Most Prolific Year: 2021 (3 Papers)
🤝 Key Collaborators: 47
🏛 Institutions: Beijing Academy of Artificial Intelligence, Harbin Institute of Technology, Shanghai Artificial Intelligence Laboratory, Northwestern Polytechnical University

Top Papers

  1. 1
  2. 2
  3. 3
  4. 4
  5. 5
  6. 6
  7. 7
  8. 8
  9. 9
  10. 10

Key Collaborators

Contact & Links

Available for collaboration
Content generated · 14 days ago