Mohammad Salman Khan

Skolkovo Institute of Science and Technology, UCSI University

Papers

2

Total Citations

5

H-Index

2

About

Mohammad Salman Khan is a robotics researcher whose work sits at the intersection of computer vision, natural language processing, and autonomous manipulation. His primary research areas include vision-language-action (VLA) models, bimanual robotic control, and human-robot interaction for service applications. Khan’s most notable contribution is the development of **Shake-VLA**, a groundbreaking Vision-Language-Action model-based system that enables bimanual robotic manipulation for automated cocktail preparation. This system integrates vision modules for ingredient detection, speech-to-text for interpreting user commands, and coordinated dual-arm control to execute complex mixing tasks—demonstrating a practical step toward general-purpose service robots. He has also contributed to sports robotics with the **Autonomous Table Tennis Ball Retrieving Robot**, which addresses the real-world challenge of collecting scattered balls during training sessions. While his citation counts are still growing (3 and 2 citations respectively for his top papers), the novelty of his VLA approach and its potential for commercial applications in hospitality and domestic settings marks him as an emerging innovator in embodied AI and robotic service systems.

Research Focus

Key Achievements

2
H-Index
2
Papers
5
Total Citations
3
Avg Citations/Paper
🏆 Most Cited Paper
Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing
3 citations · 2025
📈 Most Prolific Year: 2025 (1 Papers)
🤝 Key Collaborators: 12
🏛 Institutions: Skolkovo Institute of Science and Technology, UCSI University

Top Papers

  1. 1
  2. 2

Key Collaborators

Contact & Links

Available for collaboration
Content generated · 14 days ago