Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Ashishkumar Gudmalwar — most-cited papers & profile · Speech Audio
← authors
·
overview
Ashishkumar Gudmalwar
3
papers ·
1
citations ·
3
h-index
Sony (Taiwan) · Sony Corporation (United States)
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
DubWise: Video-Guided Speech Duration Control in Multimodal LLM-based Text-to-Speech for Dubbing
2024 · 1 citations
VECL-TTS: Voice identity and Emotional style controllable Cross-Lingual Text-to-Speech
2024
EmoReg: Directional Latent Vector Modeling for Emotional Intensity Regularization in Diffusion-based Voice Conversion
2024
Top co-authors
Nirmesh Shah
· 3
Pankaj Wasnik
· 3
Rajiv Ratn Shah
· 3
Ishan D. Biyani
· 1
Neha Sahipjohn
· 1
Sai Akarsh
· 1
Topics
Text-to-Speech
Voice Cloning
Audio Generation
Speaker Analysis
Speech Translation
Multimodal Audio
Speech Recognition