Hongxu Yin

Papers

2

Total Citations

15

H-Index

2

About

Hongxu Yin is a leading researcher at NVIDIA whose work sits at the intersection of embodied AI, robotics, and vision-language models. His primary research areas include legged robot navigation, vision-language-action models, and human-robot interaction. Yin’s most notable contribution is the development of **NaVILA** (2024–2025), a groundbreaking Vision-Language-Action model that enables legged robots to understand and execute complex natural language commands for navigation. This work addresses a critical challenge in robotics: translating human instructions like “walk forward along the way” or “stop in front of the red door” into precise, adaptive movement through cluttered and uneven terrain. The NaVILA framework has quickly gained traction, accumulating over 15 citations within its first year, signaling its importance to the field. By bridging large language models with physical robot control, Yin is helping to create robots that can follow spoken directions in real-world environments—from stepping over obstacles to navigating through doorways. His research promises to make robotic assistants more intuitive and accessible, with potential applications in search-and-rescue, home assistance, and industrial inspection.

Research Focus

Key Achievements

2
H-Index
2
Papers
15
Total Citations
8
Avg Citations/Paper
🏆 Most Cited Paper
NaVILA: Legged Robot Vision-Language-Action Model for Navigation
13 citations · 2025
📈 Most Prolific Year: 2025 (1 Papers)
🤝 Key Collaborators: 9

Top Papers

  1. 1
  2. 2

Key Collaborators

Contact & Links

Available for collaboration
Content generated · 12 days ago