Awesome Large Language Models
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
David Harwath — most-cited papers & profile · Large Language Models
← authors
·
overview
David Harwath
6
papers ·
23
citations ·
0
h-index
The University of Texas at Austin
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Transfer Learning From Audio-visual Grounding To Speech Recognition
2019 · 18 citations
VoiceStar: Robust Zero-Shot Autoregressive TTS with Duration Control and Extrapolation
2025 · 5 citations
Learning Modality-Invariant Representations for Speech and Images
2017
Text-Free Image-to-Speech Synthesis Using Learned Segmental Units
2020
Everything at Once -- Multi-modal Fusion Transformer for Video Retrieval
2021
Top co-authors
Chin-Jou Li
· 1
Daisuke Saito
· 1
David R. Mortensen
· 1
Eunjung Yeo
· 1
Jian Zhu
· 1
Kwanghee Choi
· 1
Nobuaki Minematsu
· 1
Shikhar Bharadwaj
· 1
Shinji Watanabe
· 1
Stephen McIntosh
· 1
Topics
Text-to-Speech
Audio Generation
cs.CV
Multimodal Audio
Speech Recognition
Speech Translation
eess.IV
cs.CL
cs.SD
eess.AS