Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
IDS — most-cited papers & profile · Speech Audio
← authors
·
overview
IDS
16
papers ·
5
citations
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
SpecDiff-GAN: A Spectrally-Shaped Noise Diffusion GAN for Speech and Music Synthesis
2024 · 9 citations
WaveTransfer: A Flexible End-to-end Multi-instrument Timbre Transfer with Diffusion
2024 · 3 citations
A Hybrid Model for Weakly-Supervised Speech Dereverberation
2025 · 2 citations
F-StrIPE: Fast Structure-Informed Positional Encoding for Symbolic Music Generation
2025 · 1 citations
O-EENC-SD: Efficient Online End-to-End Neural Clustering for Speaker Diarization
2025
Is Phase Really Needed for Weakly-Supervised Dereverberation ?
2025
Of All StrIPEs: Investigating Structure-informed Positional Encoding for Efficient Music Generation
2025
Unsupervised Harmonic Parameter Estimation Using Differentiable DSP and Spectral Optimal Transport
2023
Online speaker diarization of meetings guided by speech separation
2024
Structure-informed Positional Encoding for Music Generation
2024
GLA-Grad: A Griffin-Lim Extended Waveform Generation Diffusion Model
2024
A lightweight dual-stage framework for personalized speech enhancement based on DeepFilterNet2
2024
Speech dereverberation constrained on room impulse response characteristics
2024
Episodic fine-tuning prototypical networks for optimization-based few-shot learning: Application to audio classification
2024
Top co-authors
LTCI
· 7
S2A
· 7
Ga\"el Richard (S2A
· 5
Gael Richard (S2A
· 4
IP Paris
· 3
Jonathan Le Roux (MERL
· 3
Mathieu Fontaine (S2A
· 3
Teysir Baoueb (IP Paris
· 3
Changhong Wang (LTCI
· 2
Elio Gruttadauria (IP Paris
· 2
Gael Richard (IP Paris
· 2
Haocheng Liu (IP Paris
· 2
Topics
Music Generation
Audio Generation
Speech Enhancement
Audio Understanding
Speaker Analysis
Speech Translation
Speech Recognition
Text-to-Speech
Voice Cloning