Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Yang Feng — most-cited papers & profile · Speech Audio
← authors
·
overview
Yang Feng
44
papers ·
199
citations
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
DASpeech: Directed Acyclic Transformer for Fast and High-quality Speech-to-Speech Translation
2023 · 6 citations
Efficient Speech Language Modeling via Energy Distance in Continuous Latent Space
2025 · 4 citations
Memory Visualization for Gated Recurrent Neural Networks in Speech Recognition
2016 · 3 citations
Collaborative Learning for Language and Speaker Recognition
2016 · 2 citations
STEMM: Self-learning with Speech-text Manifold Mixup for Speech Translation
2022 · 2 citations
Information-Transport-based Policy for Simultaneous Translation
2022 · 2 citations
Unified Segment-to-Segment Framework for Simultaneous Sequence Generation
2023 · 2 citations
LLaMA-Omni: Seamless Speech Interaction with Large Language Models
2024 · 2 citations
LLaMA-Omni2: LLM-based Real-time Spoken Chatbot with Autoregressive Streaming Speech Synthesis
2025 · 1 citations
Understanding and Bridging the Modality Gap for Speech Translation
2023 · 1 citations
CMOT: Cross-modal Mixup via Optimal Transport for Speech Translation
2023 · 1 citations
End-to-End Simultaneous Speech Translation with Differentiable Segmentation
2023 · 1 citations
Efficient Training for Cross-lingual Speech Language Models
2026
StreamUni: Achieving Streaming Speech Translation with a Unified Large Speech-Language Model
2025
FastLongSpeech: Enhancing Large Speech-Language Models for Efficient Long-Speech Processing
2025
Top co-authors
QingKai Fang
· 14
Shaolei Zhang
· 11
Shoutao Guo
· 8
Zhengrui Ma
· 8
Yan Zhou
· 7
Min Zhang
· 6
Dong Wang
· 2
Zhiyuan Tang
· 2
Chenze Shao
· 1
Fandong Meng
· 1
Jie Zhou
· 1
Lei Li
· 1
Topics
Speech Recognition
Speech Translation
Text-to-Speech
Audio Generation
Multimodal Audio
cs.CL
Music Generation
cs.SD
Audio Understanding
cs.AI