Weibin Meng

Papers

2

Total Citations

8

H-Index

2

About

Weibin Meng is a rising researcher at the forefront of embodied AI and multimodal learning, with a focus on bridging the gap between language understanding and physical robot control. His most notable contribution is the co-authorship of **ChatVLA**, a pioneering Vision-Language-Action model that unifies multimodal perception, comprehension, and robotic manipulation into a single, coherent framework. By systematically analyzing existing VLA training paradigms, Meng and his team identified critical bottlenecks that prevent large language models from achieving holistic, real-world interaction. This work, presented at EMNLP 2025, has already garnered early attention (over 6 citations) for its promise to create robots that can both "see" and "act" with human-like intuition. Meng’s research sits at the exciting intersection of computer vision, natural language processing, and robotics, addressing fundamental questions about how machines can understand and physically engage with their environment. His work is particularly relevant for students and researchers interested in the next generation of autonomous systems, where seamless integration of perception and action is key.

Research Focus

Key Achievements

2
H-Index
2
Papers
8
Total Citations
4
Avg Citations/Paper
🏆 Most Cited Paper
ChatVLA: Unified Multimodal Understanding and Robot Control with Vision-Language-Action Model
6 citations · 2025
📈 Most Prolific Year: 2025 (2 Papers)
🤝 Key Collaborators: 11

Top Papers

  1. 1
  2. 2

Key Collaborators

Contact & Links

Available for collaboration
Content generated · 13 days ago