Vasil Khalidov

Idiap Research Institute

Papers

7

Total Citations

114

H-Index

5

About

Vasil Khalidov is a leading researcher at the intersection of human-robot interaction (HRI), multimodal perception, and self-supervised learning. His work fundamentally addresses how machines can understand and engage with humans in social, multi-party settings. Khalidov made seminal contributions to HRI by creating the **Vernissage Corpus** (2013, 33 citations), a benchmark dataset for conversational HRI that captures rich, real-behaving robot interactions with multiple humans. This work, alongside his studies on **engagement-based multi-party dialog** (2011, 28 citations) and **context-aware addressee estimation** (2013, 10 citations), established foundational methods for robots to detect visual focus of attention, recognize speakers, and decide when to engage users. His research on **finding audio-visual events** (2011, 22 citations) introduced novel multimodal clustering algorithms for detecting people who can be both seen and heard. More recently, Khalidov has pushed the frontier of AI with **V-JEPA 2** (2025, 3 citations), a self-supervised video model that learns to understand, predict, and plan by combining internet-scale video with minimal robot interaction data. This work represents a major step toward machines that can learn world models through observation, bridging perception and action. With over 100 citations across his key papers, Khalidov’s research continues to shape how robots perceive, interact with, and learn from the social world.

Research Focus

Key Achievements

5
H-Index
7
Papers
114
Total Citations
16
Avg Citations/Paper
🏆 Most Cited Paper
The vernissage corpus: A conversational Human-Robot-Interaction dataset
33 citations · 2013
📈 Most Prolific Year: 2013 (3 Papers)
🤝 Key Collaborators: 43
🏛 Institutions: Idiap Research Institute

Top Papers

  1. 1
  2. 2
  3. 3
  4. 4
  5. 5
  6. 6
  7. 7

Key Collaborators

Contact & Links

Available for collaboration
Content generated · 13 days ago