Huaijin Pi
Papers
1
Total Citations
50
H-Index
1
About
Huaijin Pi is a researcher at the forefront of embodied AI and robotic manipulation, with a primary focus on integrating vision, language, and action for intelligent grasping. His most cited work, "A Joint Modeling of Vision-Language-Action for Target-oriented Grasping in Clutter" (2023, 50 citations), addresses a critical challenge in robotics: enabling a robot to grasp a specific object based on a natural language instruction within a cluttered environment. Unlike prior approaches that separate visual grounding from grasp generation—often leading to inefficiencies and errors—Pi’s key contribution is a unified, end-to-end model that jointly processes visual and linguistic cues to directly output a targeted grasp. This innovation eliminates the need for intermediate steps, significantly improving both speed and accuracy in complex, real-world settings. By bridging the gap between high-level human commands and low-level robotic actions, Pi’s work has laid a vital foundation for more intuitive and robust human-robot interaction. His research is highly influential in the fields of robot learning and multimodal AI, demonstrating how tight integration of perception and action can unlock new capabilities for autonomous systems in homes, warehouses, and beyond.
Research Focus
Key Achievements
Top Papers
- 1