Papers

5

Total Citations

68

H-Index

4

About

Kenji Iwata is a leading researcher in multimodal scene understanding, with a focus on bridging the gap between 3D spatial perception and natural language. His core contributions lie in developing frameworks that enable robots and AI systems to recognize, describe, and reason about changes in real-world environments. Iwata pioneered the concept of scene change captioning, creating systems that generate natural language descriptions of alterations observed in indoor and multi-view 3D scenes—a critical capability for human-robot interaction and anomaly detection. His most cited work, "3D-Aware Scene Change Captioning From Multiview Images" (26 citations), demonstrates how to synthesize observations from multiple viewpoints into coherent text. He has also advanced visual question answering by introducing active viewpoint selection, allowing agents to iteratively explore scenes to answer queries. Notably, his early work on the musician robot’s speech conversation system (1985) shows a long-standing commitment to embodied AI. With over 68 citations across his top papers, Iwata’s research continues to shape how machines perceive, describe, and interact with dynamic 3D environments.

Research Focus

Key Achievements

4
H-Index
5
Papers
68
Total Citations
14
Avg Citations/Paper
🏆 Most Cited Paper
3D-Aware Scene Change Captioning From Multiview Images
26 citations · 2020
📈 Most Prolific Year: 2020 (3 Papers)
🤝 Key Collaborators: 10
🏛 Institutions: National Institute of Advanced Industrial Science and Technology, Waseda University

Top Papers

  1. 1
  2. 2
  3. 3
  4. 4
  5. 5

Key Collaborators

Contact & Links

Available for collaboration
Content generated · 14 days ago