Zhaoxi Zhang

Papers

1

Total Citations

6

H-Index

1

About

Dr. Zhaoxi Zhang is at the forefront of integrating artificial intelligence with robotic surgery, specializing in large vision-language models (LVLMs) and surgical visual question answering (Surgical-VQA). Their groundbreaking work, "Surgical-LVLM: Learning to Adapt Large Vision-Language Model for Grounded Visual Question Answering in Robotic Surgery" (2024), pioneers a novel approach that enables AI to not only answer complex questions about surgical scenes but also precisely ground those answers to specific regions in the visual field. This capability addresses a critical gap in automated surgical mentorship, moving beyond simple descriptions to provide context-aware, spatially accurate guidance. With 6 citations already in its first year, this work is rapidly gaining traction for its potential to transform personalized surgical training and intraoperative decision support. Dr. Zhang’s contributions are particularly notable for bridging the gap between general-purpose AI and the high-stakes, domain-specific demands of medicine, laying a foundation for safer, more intelligent robotic-assisted procedures. Their research stands at the intersection of computer vision, natural language processing, and surgical robotics, promising to redefine how surgeons learn and operate.

Research Focus

Key Achievements

1
H-Index
1
Papers
6
Total Citations
6
Avg Citations/Paper
🏆 Most Cited Paper
Surgical-LVLM: Learning to Adapt Large Vision-Language Model for Grounded Visual Question Answering in Robotic Surgery
6 citations · 2024
📈 Most Prolific Year: 2024 (1 Papers)
🤝 Key Collaborators: 9

Top Papers

  1. 1

Key Collaborators

Contact & Links

Available for collaboration
Content generated · 12 days ago