Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Yi-Jen Shih — most-cited papers & profile · Speech Audio
← authors
·
overview
Yi-Jen Shih
8
papers ·
3
citations ·
4
h-index
The University of Texas at Austin
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
M-SpeechCLIP: Leveraging Large-Scale, Pre-Trained Models for Multilingual Speech to Image Retrieval
2022 · 2 citations
SpeechCLIP: Integrating Speech with Pre-Trained Vision and Language Model
2022 · 1 citations
Unifying Model and Layer Fusion for Speech Foundation Models
2025
Can Speech LLMs Think while Listening?
2025
Integrating Self-supervised Speech Model with Pseudo Word-level Targets from Visually-grounded Speech Model
2024
SpeechCLIP+: Self-supervised multi-task representation learning for speech via CLIP and speech-image data
2024
Interface Design for Self-Supervised Speech Models
2024
Dynamic-SUPERB Phase-2: A Collaboratively Expanding Benchmark for Measuring the Capabilities of Spoken Language Models with 180 Tasks
2024
Top co-authors
David Harwath
· 7
Hung-yi Lee
· 5
Hsuan-Fu Wang
· 4
Layne Berry
· 4
Heng-Jui Chang
· 3
Puyuan Peng
· 3
Andy T. Liu
· 1
Ankita Pasad
· 1
Anuj Diwan
· 1
Chao-Han Huck Yang
· 1
Chee-En Yu
· 1
Chen-An Li
· 1
Topics
Speech Recognition
Audio Understanding
Multimodal Audio
Text-to-Speech
Speech Translation
Speaker Analysis