Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Jiatong Shi — most-cited papers & profile · Speech Audio
← authors
·
overview
Jiatong Shi
39
papers ·
164
citations ·
0
h-index
Carnegie Mellon University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
SUPERB: Speech processing Universal PERformance Benchmark
2021 · 51 citations
Recent Developments on ESPnet Toolkit Boosted by Conformer
2020 · 40 citations
ESPnet2-TTS: Extending the Edge of TTS Research
2021 · 28 citations
UniAudio: An Audio Foundation Model Toward Universal Audio Generation
2023 · 14 citations
SVDD 2024: The Inaugural Singing Voice Deepfake Detection Challenge
2024 · 11 citations
SVDD Challenge 2024: A Singing Voice Deepfake Detection Challenge Evaluation Plan
2024 · 3 citations
SingMOS: An extensive Open-Source Singing Voice Dataset for MOS Prediction
2024 · 3 citations
Exploring Speech Recognition, Translation, and Understanding with Discrete Speech Units: A Comparative Study
2023 · 2 citations
EFFUSE: Efficient Self-Supervised Feature Fusion for E2E ASR in Low Resource and Multilingual Scenarios
2023 · 2 citations
OWSM v3.1: Better and Faster Open Whisper-Style Speech Models based on E-Branchformer
2024 · 2 citations
Joint Beam Search Integrating CTC, Attention, and Transducer Decoders
2024 · 2 citations
ESPnet-SpeechLM: An Open Speech Language Model Toolkit
2025 · 1 citations
How to Learn a New Language? An Efficient Solution for Self-Supervised Learning Models Unseen Languages Adaption in Low-Resource Scenario
2024 · 1 citations
ESPnet-ST IWSLT 2021 Offline Speech Translation System
2021 · 1 citations
Cross-lingual Transfer for Speech Processing using Acoustic Language Similarity
2021 · 1 citations
Top co-authors
Shinji Watanabe
· 26
Jinchuan Tian
· 13
Xuankai Chang
· 10
Yuxun Tang
· 10
Hung-yi Lee
· 7
Siddhant Arora
· 7
Jionghao Han
· 6
Wangyou Zhang
· 5
Brian Yan
· 4
Jee-weon Jung
· 4
Karen Livescu
· 4
William Chen
· 4
Topics
Speech Recognition
Audio Understanding
Audio Generation
Music Generation
Text-to-Speech
Speech Translation
Speech Enhancement
Speaker Analysis
Multimodal Audio
Voice Cloning