Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Heinrich Dinkel — most-cited papers & profile · Multimodal
← authors
·
overview
Heinrich Dinkel
5
papers ·
73
citations ·
18
h-index
PLA Academy of Military Science · Xiaomi (China)
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Voice activity detection in the wild: A data-driven approach using teacher-student training
2021 · 34 citations
MiDashengLM: Efficient Audio Understanding with General Audio Captions
2025 · 30 citations
GLAP: General contrastive audio-text pretraining across domains and languages
2025 · 9 citations
An empirical study of weakly supervised audio tagging embeddings for general audio representations
2022
Topics
Audio Understanding
Speech Recognition
Multimodal Audio
Audio Generation
Voice Cloning
Music Generation