Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Bin Ma — most-cited papers & profile · Speech Audio
← authors
·
overview
Bin Ma
38
papers ·
259
citations ·
0
h-index
Yili Normal University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
MinMo: A Multimodal Large Language Model for Seamless Voice Interaction
2025 · 86 citations
MossFormer: Pushing the Performance Limit of Monaural Speech Separation using Gated Single-Head Transformer with Convolution-Augmented Joint Self-Attentions
2023 · 64 citations
Monaural Speech Enhancement with Complex Convolutional Block Attention Module and Joint Time Frequency Losses
2021 · 49 citations
D2Former: A Fully Complex Dual-Path Dual-Decoder Conformer Network using Joint Complex Masking and Complex Spectral Mapping for Monaural Speech Enhancement
2023 · 16 citations
Towards Natural and Controllable Cross-Lingual Voice Conversion Based on Neural TTS Model and Phonetic Posteriorgram
2021 · 13 citations
M2MeT: The ICASSP 2022 Multi-Channel Multi-Party Meeting Transcription Challenge
2021 · 11 citations
Independent language modeling architecture for end-to-end ASR
2019 · 6 citations
Learning Acoustic Word Embeddings with Temporal Context for Query-by-Example Speech Search
2018 · 4 citations
Towards Natural Bilingual and Code-Switched Speech Synthesis Based on Mix of Monolingual Recordings and Cross-Lingual Voice Conversion
2020 · 3 citations
Conditional Latent Diffusion-Based Speech Enhancement Via Dual Context Learning
2025 · 1 citations
FRCRN: Boosting Feature Representation using Frequency Recurrence for Monaural Speech Enhancement
2022 · 1 citations
Adaptive Knowledge Distillation between Text and Speech Pre-trained Models
2023 · 1 citations
Contrastive Speech Mixup for Low-resource Keyword Spotting
2023 · 1 citations
MossFormer2: Combining Transformer and RNN-Free Recurrent Network for Enhanced Time-Domain Monaural Speech Separation
2023 · 1 citations
Speech Separation using Neural Audio Codecs with Embedding Loss
2024 · 1 citations
Top co-authors
Chongjia Ni
· 14
Chong Zhang
· 13
Yukun Ma
· 12
Dianwen Ng
· 10
Eng Siong Chng
· 10
Hao Wang
· 9
Kun Zhou
· 9
Zhihao Du
· 5
Wen Wang
· 4
Fan Yu
· 3
Lei Xie
· 3
Qian Chen
· 3
Topics
Speech Recognition
Speech Enhancement
Audio Understanding
Text-to-Speech
Speech Translation
Audio Generation
Multimodal Audio
Speaker Analysis
Voice Cloning
eess.AS