Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Roshan Sharma — most-cited papers & profile · Speech Audio
← authors
·
overview
Roshan Sharma
10
papers ·
7
citations ·
13
h-index
Texas Tech University · University of Oklahoma
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Augsumm: Towards Generalizable Speech Summarization Using Synthetic Labels From Large Language Model
2024 · 2 citations
XNOR-FORMER: Learning Accurate Approximations in Long Speech Transformers
2022 · 2 citations
Exploring Speech Recognition, Translation, and Understanding with Discrete Speech Units: A Comparative Study
2023 · 2 citations
UniverSLU: Universal Spoken Language Understanding for Diverse Tasks with Natural Language Instructions
2023 · 1 citations
SLUE Phase-2: A Benchmark Suite of Diverse Spoken Language Understanding Tasks
2022
BASS: Block-wise Adaptation for Speech Summarization
2023
Dynamic-SUPERB: Towards A Dynamic, Collaborative, and Comprehensive Instruction-Tuning Benchmark for Speech
2023
Reproducing Whisper-Style Training Using an Open-Source Toolkit and Publicly Available Data
2023
Evaluating Speech Synthesis by Training Recognizers on Synthetic Speech
2023
On the Evaluation of Speech Foundation Models for Spoken Language Understanding
2024
Top co-authors
Shinji Watanabe
· 7
Siddhant Arora
· 6
Jee-weon Jung
· 5
Bhiksha Raj
· 4
Hung-yi Lee
· 3
Soumi Maiti
· 3
Yifan Peng
· 3
Ankita Pasad
· 2
Hira Dhamyal
· 2
Jiatong Shi
· 2
Jinchuan Tian
· 2
Karen Livescu
· 2
Topics
Speech Recognition
Audio Understanding
Text-to-Speech
Speech Translation
cs.CL
Multimodal Audio