Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Bin Wang — most-cited papers & profile · Speech Audio
← authors
·
overview
Bin Wang
110
papers ·
677
citations ·
0
h-index
Binzhou University · Southeast University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Language modeling with Neural trans-dimensional random fields
2017 · 12 citations
Attention-based Transducer for Online Speech Recognition
2020 · 6 citations
Model Interpolation with Trans-dimensional Random Field Language Models for Speech Recognition
2016 · 4 citations
AudioBench: A Universal Benchmark for Audio Large Language Models
2024 · 4 citations
Attention-based sequence-to-sequence model for speech recognition: development of state-of-the-art system on LibriSpeech and its application to non-native English
2018 · 2 citations
SACodec: Asymmetric Quantization with Semantic Anchoring for Low-Bitrate High-Fidelity Neural Speech Codecs
2025
Attention2Probability: Attention-Driven Terminology Probability Estimation for Robust Speech-to-Text System
2025
Step-Audio 2 Technical Report
2025
Step-Audio-AQAA: a Fully End-to-End Expressive Large Audio Language Model
2025
NTU Speechlab LLM-Based Multilingual ASR System for Interspeech MLC-SLM Challenge 2025
2025
Joint Training And Decoding for Multilingual End-to-End Simultaneous Speech Translation
2025
Step-Audio: Unified Understanding and Generation in Intelligent Speech Interaction
2025
Learning neural trans-dimensional random field language models with noise-contrastive estimation
2017
The Volcspeech system for the ICASSP 2022 multi-channel multi-party meeting transcription challenge
2022
Streaming Audio Transformers for Online Audio Tagging
2023
Top co-authors
Bingxin Li
· 3
Binxing Jiao
· 3
Bo Li
· 3
Changxin Miao
· 3
Changyi Wan
· 3
Chao Yan
· 3
Chen Hu
· 3
Chen Xu
· 3
Dapeng Shi
· 3
Daxin Jiang
· 3
Dingyuan Hu
· 3
Enle Liu
· 3
Topics
Speech Recognition
Speech Translation
Audio Understanding
Multimodal Audio
cs.SD
cs.CL
Audio Generation
Text-to-Speech
eess.AS
Music Generation