Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Subhankar Ghosh — most-cited papers & profile · Speech Audio
← authors
·
overview
Subhankar Ghosh
11
papers ·
11
citations ·
3
h-index
University of Technology Sydney
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Improving Robustness of LLM-based Speech Synthesis by Learning Monotonic Alignment
2024 · 8 citations
VANI: Very-lightweight Accent-controllable TTS for Native and Non-native speakers with Identity Preservation
2023 · 2 citations
Koel-TTS: Enhancing LLM based Speech Generation with Preference Alignment and Classifier Free Guidance
2025 · 1 citations
NanoCodec: Towards High-Quality Ultra Fast Speech LLM Inference
2025
SALM-Duplex: Efficient and Direct Duplex Modeling for Speech-to-Speech Language Model
2025
Adapter-Based Extension of Multi-Speaker Text-to-Speech Model for New Speakers
2022
SALM: Speech-augmented Language Model with In-context Learning for Speech Recognition and Translation
2023
Low Frame-rate Speech Codec: a Codec Designed for Fast High-quality Speech LLM Training and Inference
2024
Top co-authors
Boris Ginsburg
· 6
Jason Li
· 6
Edresson Casanova
· 4
Paarth Neekhara
· 4
Shehzeen Hussain
· 4
Rafael Valle
· 3
Ante Juki\'c
· 2
Jagadeesh Balam
· 2
Rohan Badlani
· 2
Ryan Langman
· 2
Xuesong Yang
· 2
Zhehuai Chen
· 2
Topics
Text-to-Speech
Audio Generation
Speech Recognition
Speech Translation
Multimodal Audio
Voice Cloning
Music Generation
Audio Understanding
Speech Enhancement