Zongxia Li

University of Maryland, College Park

Papers

1

Total Citations

18

H-Index

1

About

Zongxia Li is a rising researcher at the forefront of multimodal artificial intelligence, with a primary focus on large vision-language models (VLMs) and their real-world applications. Her most-cited work, a comprehensive 2025 survey on the benchmark evaluations, applications, and challenges of large vision-language models, has already garnered 18 citations, reflecting the timely importance of her contributions. In this survey, Li systematically analyzes how models like CLIP and Claude bridge computer vision and natural language processing, enabling machines to perceive and reason through both visual and textual modalities. Her work not only catalogs existing benchmarks but also identifies critical gaps in evaluation methodologies, offering a roadmap for future research. By synthesizing the transformative potential of VLMs—from image-text retrieval to complex reasoning tasks—Li provides an essential resource for students and researchers navigating this rapidly evolving field. Her contributions are particularly notable for their clarity and depth, making complex multimodal systems accessible to a broader audience. As a scholar dedicated to advancing AI’s perceptual capabilities, Zongxia Li is establishing herself as a key voice in the next generation of multimodal AI research.

Research Focus

Key Achievements

1
H-Index
1
Papers
18
Total Citations
18
Avg Citations/Paper
🏆 Most Cited Paper
Benchmark Evaluations, Applications, and Challenges of Large Vision Language Models: A Survey
18 citations · 2025
📈 Most Prolific Year: 2025 (1 Papers)
🤝 Key Collaborators: 4
🏛 Institutions: University of Maryland, College Park

Top Papers

  1. 1

Key Collaborators

Contact & Links

Available for collaboration
Content generated · 12 days ago