Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Toshimitsu Uesaka — most-cited papers & profile · Speech Audio
← authors
·
overview
Toshimitsu Uesaka
11
papers ·
35
citations ·
7
h-index
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
SQ-VAE: Variational Bayes on Discrete Representation with Self-annealed Stochastic Quantization
2022 · 13 citations
Diffiner: A Versatile Diffusion-based Generative Refiner for Speech Enhancement
2022 · 1 citations
DiffRoll: Diffusion-based Generative Music Transcription with Unsupervised Pretraining Capability
2022
Top co-authors
Naoki Murata
· 3
Shusuke Takahashi
· 3
Yuki Mitsufuji
· 3
Ryosuke Sawata
· 2
Takashi Shibuya
· 2
Yuhta Takida
· 2
Chieh-Hsin Lai
· 1
Dorien Herremans
· 1
Junki Ohmura
· 1
Kin Wai Cheuk
· 1
Naoya Takahashi
· 1
Toshiyuki Kumakura
· 1
Topics
Audio Generation
cs.CV
cs.LG
eess.IV
Music Generation
Audio Understanding
Speech Enhancement