Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Emiru Tsunoo — most-cited papers & profile · Speech Audio
← authors
·
overview
Emiru Tsunoo
30
papers ·
39
citations ·
14
h-index
Sony Corporation (United States)
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Towards Online End-to-end Transformer Automatic Speech Recognition
2019 · 31 citations
Differentiable K-means for Fully-optimized Discrete Token-based ASR
2025 · 2 citations
Scheduled Interleaved Speech-Text Training for Speech-to-Speech Translation with LLMs
2025 · 1 citations
Transformer ASR with Contextual Block Processing
2019 · 1 citations
Streaming Transformer ASR with Blockwise Synchronous Beam Search
2020 · 1 citations
A Study on the Integration of Pipeline and E2E SLU systems for Spoken Semantic Parsing toward STOP Quality Challenge
2023 · 1 citations
Decoder-only Architecture for Speech Recognition with CTC Prompts and Text Data Augmentation
2023 · 1 citations
UniverSLU: Universal Spoken Language Understanding for Diverse Tasks with Natural Language Instructions
2023 · 1 citations
Benchmarking Speech-to-Speech Translation Models
2026
DialogueSidon: Recovering Full-Duplex Dialogue Tracks from In-the-Wild Dialogue Audio
2026
Phonological Tokenizer: Prosody-Aware Phonetic Token via Multi-Objective Fine-Tuning with Differentiable K-Means
2026
Chain-of-Thought Training for Open E2E Spoken Dialogue Systems
2025
Whale: Large-Scale multilingual ASR model with w2v-BERT and E-Branchformer with large speech data
2025
End-to-end Adaptation with Backpropagation through WFST for On-device Speech Recognition System
2019
Gaussian Kernelized Self-Attention for Long Sequence Data and Its Application to CTC-based Speech Recognition
2021
Top co-authors
Yosuke Kashiwagi
· 27
Shinji Watanabe
· 23
Hayato Futami
· 19
Siddhant Arora
· 15
Chaitanya Narisetty
· 4
Toshiyuki Kumakura
· 4
Yifan Peng
· 3
Jee-weon Jung
· 2
Jessica Huynh
· 2
Kentaro Onda
· 2
Michael Hentschel
· 2
Satoshi Asakawa
· 2
Topics
Speech Recognition
Speech Translation
Text-to-Speech
Audio Understanding
Speech Enhancement
Audio Generation
Multimodal Audio
Music Generation
Speaker Analysis