Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Xavier Serra — most-cited papers & profile · Speech Audio
← authors
·
overview
Xavier Serra
16
papers ·
108
citations ·
44
h-index
Universitat Pompeu Fabra · Escola Superior de Música de Catalunya · Hospital de Sabadell
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
musicnn: Pre-trained convolutional neural networks for music audio tagging
2019 · 41 citations
Toward Interpretable Music Tagging with Self-Attention
2019 · 35 citations
Semi-Supervised Music Tagging Transformer
2021 · 13 citations
COALA: Co-Aligned Autoencoders for Learning Semantically Enriched Audio Representations
2020 · 9 citations
Score-informed syllable segmentation for a cappella singing voice with convolutional neural networks
2017 · 8 citations
Metrical-accent Aware Vocal Onset Detection in Polyphonic Audio
2017 · 2 citations
TIV.lib: an open-source library for the tonal description of musical audio
2020 · 2 citations
End-to-end music source separation: is it possible in the waveform domain?
2018 · 1 citations
TensorFlow Audio Models in Essentia
2020 · 1 citations
Matching Text and Audio Embeddings: Exploring Transfer-learning Strategies for Language-based Audio Retrieval
2022 · 1 citations
An objective evaluation of Hearing Aids and DNN-based speech enhancement in complex acoustic scenes
2023 · 1 citations
Efficient and Fast Generative-Based Singing Voice Separation using a Latent Diffusion Model
2025
Generating Separated Singing Vocals Using a Diffusion Model Conditioned on Music Mixtures
2025
Learning Contextual Tag Embeddings for Cross-Modal Alignment of Audio and Tags
2020
Leveraging Pre-Trained Autoencoders for Interpretable Prototype Learning of Music Audio
2024
Top co-authors
Jordi Pons
· 4
Dmitry Bogdanov
· 2
Gen\'is Plaja-Roglans
· 2
Igor Pereira
· 2
Konstantinos Drossos
· 2
Minz Won
· 2
Tuomas Virtanen
· 2
Xavier Favory
· 2
Yun-Ning Hung
· 2
Ajay Srinivasamurthy
· 1
Andre Holzapfel
· 1
António Ramires
· 1
Topics
Audio Understanding
Music Generation
Multimodal Audio
Speech Recognition
cs.SD
cs.AI
Audio Generation
cs.IR
cs.SC
Speech Enhancement