Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Daisuke Niizumi — most-cited papers & profile · Speech Audio
← authors
·
overview
Daisuke Niizumi
6
papers ·
11
citations ·
12
h-index
NTT (United States) · Tokyo Metropolitan University · Intuit (United States)
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Composing General Audio Representation by Fusing Multilayer Features of a Pre-trained Model
2022 · 11 citations
Rethinking Masking Strategies for Masked Prediction-based Audio Self-supervised Learning
2026
CLAP-ART: Automated Audio Captioning with Semantic-rich Audio Representation Tokenizer
2025
Introducing Auxiliary Text Query-modifier to Content-based Audio Retrieval
2022
Masked Modeling Duo for Speech: Specializing General-Purpose Audio Representation to Speech using Denoising Distillation
2023
M2D-CLAP: Masked Modeling Duo Meets CLAP for Learning General-purpose Audio-Language Representation
2024
Top co-authors
Daiki Takeuchi
· 5
Noboru Harada
· 5
Yasunori Ohishi
· 5
and Kunio Kashino
· 2
Masahiro Yasuda
· 2
and Keisuke Imoto
· 1
and Nobutaka Ono
· 1
Binh Thien Nguyen
· 1
Binh Thien Nguyen
· 1
Daiki Takeuchi
· 1
Kunio Kashino
· 1
Masahiro Yasuda
· 1
Topics
Audio Understanding
Multimodal Audio
Speech Recognition
Audio Generation
Speech Enhancement