Shijun Xiao
Papers
1
Total Citations
2
H-Index
1
About
Dr. Shijun Xiao is a leading researcher at the intersection of human-robot interaction (HRI) and multimodal artificial intelligence, with a core focus on enabling more intuitive and natural communication between humans and machines. His most notable contribution is the development of the Visual-Speech-Text Large Language Model for HRI (VST-LLM HRI), a groundbreaking framework that integrates visual, speech, and textual inputs to create a closed-loop system for robot perception, task planning, and control. This work, published in 2025, demonstrates how large language models can be leveraged as powerful reasoning engines without requiring fine-tuning, significantly advancing the field of multimodal robotics. With over 2 citations already, Dr. Xiao’s research is gaining rapid recognition for its practical implications in assistive robotics and autonomous systems. His innovative design of a Modality Language Model (MLM) bridges the gap between diverse sensory data and actionable robot commands, offering a scalable solution for real-world HRI challenges. Dr. Xiao’s work is essential reading for students and researchers interested in the future of embodied AI and seamless human-machine collaboration.
Research Focus
Key Achievements
Top Papers
- 1