Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Lu Lu — most-cited papers & profile · Speech Audio
← authors
·
overview
Lu Lu
46
papers ·
58
citations ·
0
h-index
Central South University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
SALMONN-omni: A Standalone Speech LLM without Codec Injection for Full-duplex Conversation
2025 · 16 citations
Seed LiveInterpret 2.0: End-to-end Simultaneous Speech-to-speech Translation with Your Voice
2025 · 6 citations
SA-SOT: Speaker-Aware Serialized Output Training for Multi-Talker ASR
2024 · 6 citations
PolyVoice: Language Models for Speech to Speech Translation
2023 · 5 citations
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
2024 · 5 citations
Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition
2024 · 5 citations
Language-specific Acoustic Boundary Learning for Mandarin-English Code-switching Speech Recognition
2023 · 2 citations
A unified multichannel far-field speech recognition system: combining neural beamforming with attention based end-to-end model
2024 · 1 citations
Enabling Auditory Large Language Models for Automatic Speech Quality Evaluation
2024 · 1 citations
SALMONN-omni: A Codec-free LLM for Full-duplex Speech Understanding and Generation
2024 · 1 citations
ParaS2S: Benchmarking and Aligning Spoken Language Models for Paralinguistic-aware Speech-to-Speech Interaction
2025
Solla: Towards a Speech-Oriented LLM That Hears Acoustic Context
2025
A Comprehensive Solution To Connect Speech Encoder And Large Language Model For ASR
2024
Spatial Attention for Far-field Speech Recognition with Deep Beamforming Neural Networks
2019
Random Utterance Concatenation Based Data Augmentation for Improving Short-video Speech Recognition
2022
Top co-authors
Jun Zhang
· 13
Yuxuan Wang
· 13
Zejun Ma
· 7
Chao Zhang
· 4
Guangzhi Sun
· 4
Wei Li
· 4
Xiaohai Tian
· 4
Chen Shen
· 3
Siyin Wang
· 3
Tao Han
· 3
Ye Bai
· 3
Yuping Wang
· 3
Topics
Speech Recognition
Text-to-Speech
Audio Understanding
Speech Translation
Speech Enhancement
Multimodal Audio
Audio Generation
eess.AS
Speaker Analysis
cs.CL