Shivam Akhouri

Vellore Institute of Technology University

Papers

1

Total Citations

4

H-Index

1

About

Shivam Akhouri is an emerging researcher at the intersection of speech processing and artificial intelligence, with a primary focus on emotion-aware speech synthesis and recognition. His most notable contribution, "EmoSRE: Emotion prediction based speech synthesis and refined speech recognition using large language model and prosody encoding" (2025), introduces a novel framework that integrates large language models with prosody encoding to simultaneously predict emotional states from speech and enhance recognition accuracy. This work, already garnering 4 citations shortly after publication, demonstrates his ability to bridge affective computing with practical speech technologies. Akhouri’s research addresses critical challenges in human-computer interaction, aiming to make synthetic speech more natural and responsive to emotional context. While early in his career, his innovative approach to combining LLMs with prosodic features signals a promising trajectory in advancing emotionally intelligent speech systems. His work holds potential for applications in virtual assistants, mental health monitoring, and adaptive communication tools, marking him as a rising voice in the field of speech AI.

Research Focus

Key Achievements

1
H-Index
1
Papers
4
Total Citations
4
Avg Citations/Paper
🏆 Most Cited Paper
EmoSRE: Emotion prediction based speech synthesis and refined speech recognition using large language model and prosody encoding
4 citations · 2025
📈 Most Prolific Year: 2025 (1 Papers)
🤝 Key Collaborators: 1
🏛 Institutions: Vellore Institute of Technology University

Top Papers

  1. 1

Key Collaborators

Contact & Links

Available for collaboration
Content generated · 12 days ago