Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Jee-weon Jung — most-cited papers & profile · Speech Audio
← authors
·
overview
Jee-weon Jung
31
papers ·
204
citations ·
21
h-index
Carnegie Mellon University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
RawNet: Advanced end-to-end deep neural network using raw waveforms for text-independent speaker verification
2019 · 162 citations
Improving Design of Input Condition Invariant Speech Enhancement
2024 · 9 citations
Improved RawNet with Feature Map Scaling for Text-independent Speaker Verification using Raw Waveforms
2020 · 7 citations
ESPnet-Codec: Comprehensive Training and Evaluation of Neural Codecs for Audio, Music, and Speech
2024 · 5 citations
Large-scale learning of generalised representations for speaker recognition
2022 · 4 citations
Pushing the limits of raw waveform speaker recognition
2022 · 3 citations
Augsumm: Towards Generalizable Speech Summarization Using Synthetic Labels From Large Language Model
2024 · 2 citations
Exploring Speech Recognition, Translation, and Understanding with Discrete Speech Units: A Comparative Study
2023 · 2 citations
One model to rule them all ? Towards End-to-End Joint Speaker Diarization and Speech Recognition
2023 · 2 citations
OWSM v3.1: Better and Faster Open Whisper-Style Speech Models based on E-Branchformer
2024 · 2 citations
Segment Aggregation for short utterances speaker verification using raw waveforms
2020 · 1 citations
Advancing the dimensionality reduction of speaker embeddings for speaker diarisation: disentangling noise and informing speech activity
2021 · 1 citations
High-resolution embedding extractor for speaker diarisation
2022 · 1 citations
UniverSLU: Universal Spoken Language Understanding for Diverse Tasks with Natural Language Instructions
2023 · 1 citations
TMT: Tri-Modal Translation between Speech, Image, and Text by Processing Different Modalities as Different Languages
2024 · 1 citations
Top co-authors
Shinji Watanabe
· 20
Hee-Soo Heo
· 8
Jinchuan Tian
· 8
Joon Son Chung
· 8
Bong-Jin Lee
· 7
Siddhant Arora
· 6
Soumi Maiti
· 6
Wangyou Zhang
· 6
Youngki Kwon
· 6
Roshan Sharma
· 5
Xuankai Chang
· 5
Jiatong Shi
· 4
Topics
Audio Understanding
Speech Recognition
Speaker Analysis
Text-to-Speech
Speech Translation
Speech Enhancement
Multimodal Audio
Audio Generation
cs.CL
Music Generation