Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Tuomas Virtanen — most-cited papers & profile · Speech Audio
← authors
·
overview
Tuomas Virtanen
23
papers ·
165
citations ·
0
h-index
Åbo Akademi University · Tampere University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Sound Event Detection in Multichannel Audio Using Spatial and Harmonic Features
2017 · 87 citations
Stacked Convolutional and Recurrent Neural Networks for Music Emotion Recognition
2017 · 45 citations
COALA: Co-Aligned Autoencoders for Learning Semantically Enriched Audio Representations
2020 · 9 citations
Deep neural network Based Low-latency Speech Separation with Asymmetric analysis-Synthesis Window Pair
2021 · 8 citations
Sound event detection via dilated convolutional recurrent neural networks
2019 · 5 citations
Multi-channel Replay Speech Detection using an Adaptive Learnable Beamformer
2025 · 4 citations
Deep neural network based speech separation optimizing an objective estimator of intelligibility for low latency applications
2018 · 2 citations
Zero-Shot Audio Classification with Factored Linear and Nonlinear Acoustic-Semantic Projections
2020 · 2 citations
Attractor-Based Speech Separation of Multiple Utterances by Unknown Number of Speakers
2025 · 1 citations
Simultaneous or Sequential Training? How Speech Representations Cooperate in a Multi-Task Self-Supervised Learning System
2023 · 1 citations
Inter-Speaker Relative Cues for Text-Guided Target Speech Extraction
2025
Sound Event Detection Using Spatial Features and Convolutional Recurrent Neural Network
2017
A Recurrent Encoder-Decoder Approach with Skip-filtering Connections for Monaural Singing Voice Separation
2017
End-to-End Polyphonic Sound Event Detection Using Convolutional Recurrent Neural Networks with Learned Time-Frequency Representation Input
2018
Low-Latency Deep Clustering For Speech Separation
2019
Top co-authors
Konstantinos Drossos
· 7
Archontis Politis
· 5
Gaurav Naithani
· 4
Sharath Adavanne
· 3
Huang Xie
· 2
Okko R\"as\"anen
· 2
Pasi Pertil\"a
· 2
Wang Dai
· 2
Xavier Favory
· 2
Xavier Serra
· 2
Yuzhu Wang
· 2
Dasa Ticha
· 1
Topics
Audio Understanding
Speech Enhancement
Speech Recognition
Speech Translation
Speaker Analysis
Multimodal Audio
stat.ML
eess.SP
Text-to-Speech
cs.NE