Awesome Large Language Models
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Long Ma — most-cited papers & profile · Large Language Models
← authors
·
overview
Long Ma
31
papers ·
64
citations ·
26
h-index
Anhui University · Anhui University of Technology
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Multi-head Monotonic Chunkwise Attention For Online Speech Recognition
2020 · 13 citations
Improving Streaming Transformer Based ASR Under a Framework of Self-supervised Learning
2021 · 13 citations
Improving Speech Recognition Accuracy of Local POI Using Geographical Models
2021 · 6 citations
Improving Hybrid CTC/Attention End-to-end Speech Recognition with Pretrained Acoustic and Language Model
2021 · 6 citations
Efficient Conformer with Prob-Sparse Attention Mechanism for End-to-EndSpeech Recognition
2021 · 5 citations
Improving Accent Identification and Accented Speech Recognition Under a Framework of Self-supervised Learning
2021 · 4 citations
MPE-TTS: Customized Emotion Zero-Shot Text-To-Speech Using Multi-Modal Prompt
2025 · 3 citations
VITA-1.5: Towards GPT-4o Level Real-Time Vision and Speech Interaction
2025 · 3 citations
Freeze-Omni: A Smart and Low Latency Speech-to-speech Dialogue Model with Frozen LLM
2024 · 3 citations
DiffCSS: Diverse and Expressive Conversational Speech Synthesis with Diffusion Models
2025 · 2 citations
Tiny Transducer: A Highly-efficient Speech Recognition Model on Edge Devices
2021 · 2 citations
Kimi K2.5: Visual Agentic Intelligence
2026 · 1 citations
Kimi K2.5: Visual Agentic Intelligence
2026 · 1 citations
DistillW2V2: A Small and Streaming Wav2vec 2.0 Based ASR Model
2023 · 1 citations
A Transcription Prompt-based Efficient Audio Large Language Model for Robust Speech Recognition
2024 · 1 citations
Topics
Speech Recognition
Speech Translation
Text-to-Speech
Multimodal Audio
Audio Generation
Speech Enhancement
Image Restoration
Audio Understanding
Multi-Agent
Orchestration