Fangyun Wei

Microsoft Research Asia (China)

Papers

2

Total Citations

7

H-Index

2

About

Fangyun Wei is a leading researcher at the intersection of computer vision, natural language processing, and robotics, with a core focus on developing foundational models for robotic manipulation. His work is pioneering the integration of large vision-language models (VLMs) into embodied AI, creating systems that can understand language commands and execute complex physical tasks. Wei’s major contributions include the development of CogACT, a foundational Vision-Language-Action (VLA) model that synergizes cognition and action, enabling robots to generalize to unseen scenarios with remarkable proficiency. He also introduced UniGraspTransformer, a universal Transformer-based network that dramatically simplifies the training pipeline for dexterous robotic grasping, moving beyond complex, multi-step methods like UniDexGrasp++. His research is highly influential, with his most recent works already garnering significant attention and citations within the AI and robotics communities. Wei’s achievements are shaping the next generation of intelligent, language-guided robotic systems, making him a key figure to watch in the rapidly evolving field of embodied AI.

Research Focus

Key Achievements

2
H-Index
2
Papers
7
Total Citations
4
Avg Citations/Paper
🏆 Most Cited Paper
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
5 citations · 2024
📈 Most Prolific Year: 2024 (1 Papers)
🤝 Key Collaborators: 25
🏛 Institutions: Microsoft Research Asia (China)

Top Papers

  1. 1
  2. 2

Key Collaborators

Contact & Links

Available for collaboration
Content generated · 13 days ago