Jinghuan Shang

Stony Brook University

Papers

5

Total Citations

37

H-Index

3

About

Jinghuan Shang is a rising star in robot learning, whose work lies at the intersection of imitation learning, reinforcement learning (RL), and vision-language models. He is best known for pioneering **third-person imitation learning (TPIL)** , a paradigm that allows robots to learn action policies by observing human demonstrations from a third-person perspective—eliminating the costly need for first-person robot data. His foundational paper on self-supervised disentangled representations for TPIL (2021, 14 citations) has become a key reference in the field. Shang also introduced **StARformer** (2022, 10 citations), a Transformer architecture that models state-action-reward sequences for more sample-efficient robot RL. His critical investigation into whether self-supervised learning truly benefits RL from pixels (2022, 9 citations) has helped shape best practices in the community. Most recently, with **LLaRA** (2024), Shang is pushing the frontier of Vision-Language-Action models, showing how to supercharge limited robot data to adapt pretrained VLMs for robotic control. Across these contributions, Shang’s work consistently addresses the data-efficiency bottleneck in robot learning, making him a leading voice in developing more practical, generalizable robotic systems.

Research Focus

Key Achievements

3
H-Index
5
Papers
37
Total Citations
7
Avg Citations/Paper
🏆 Most Cited Paper
Self-Supervised Disentangled Representation Learning for Third-Person Imitation Learning
14 citations · 2021
📈 Most Prolific Year: 2021 (2 Papers)
🤝 Key Collaborators: 13
🏛 Institutions: Stony Brook University

Top Papers

  1. 1
  2. 2
  3. 3
  4. 4
  5. 5

Key Collaborators

Contact & Links

Available for collaboration
Content generated · 13 days ago