Naoto Tsukamoto
Papers
3
Total Citations
14
H-Index
3
About
Naoto Tsukamoto is a robotics researcher advancing the frontier of autonomous systems through the integration of vision-language models (VLMs) and human-robot interaction. His work centers on enabling mobile robots to perceive and act in dynamic, unstructured environments without prior mapping or extensive training. In his highly cited 2023 paper, Tsukamoto introduced a method for semantic scene difference detection during daily-life patrolling, leveraging pre-trained large-scale VLMs to identify environmental changes—a critical capability for domestic service robots. This approach outperformed traditional anomaly detection techniques by focusing on semantic understanding rather than pixel-level differences. Building on this, his 2024 work on reflex-based open-vocabulary navigation demonstrated how an omnidirectional camera paired with multiple VLMs allows robots to navigate and respond to natural language commands without SLAM or reinforcement learning, achieving robust performance in zero-shot scenarios. Tsukamoto also developed a chat-based system for teaching robot action instructions, bridging the gap between non-expert users and complex robotic control. With over 14 citations across his recent papers, Tsukamoto’s contributions are shaping a future where robots understand and adapt to human environments intuitively, making him a rising voice in embodied AI and service robotics.
Research Focus
Key Achievements
Top Papers
- 1
- 2
- 3