Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Jade Copet — most-cited papers & profile · Speech Audio
← authors
·
overview
Jade Copet
27
papers ·
1060
citations ·
14
h-index
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
High Fidelity Neural Audio Compression
2022 · 280 citations
Speech Resynthesis from Discrete Disentangled Self-Supervised Representations
2021 · 206 citations
Simple and Controllable Music Generation
2023 · 66 citations
Textless Speech Emotion Conversion Using Discrete And Decomposed Representations
2021 · 26 citations
Text-Free Prosody-Aware Generative Spoken Language Modeling
2021 · 16 citations
Brouhaha: multi-task training for voice activity detection, speech-to-noise ratio, and C50 room acoustics estimation
2022 · 8 citations
Textually Pretrained Speech Language Models
2023 · 7 citations
Generative Spoken Dialogue Language Modeling
2022 · 3 citations
Augmentation Invariant Discrete Representation for Generative Spoken Language Modeling
2022 · 2 citations
Masked Audio Generation using a Single Non-Autoregressive Transformer
2024 · 2 citations
ASR4REAL: An extended benchmark for speech models
2021 · 1 citations
Audio Language Modeling using Perceptually-Guided Discrete Representations
2022 · 1 citations
Pushing the performances of ASR models on English and Spanish accents
2022 · 1 citations
Low-Resource Self-Supervised Learning with SSL-Enhanced TTS
2023 · 1 citations
Generative Spoken Language Modeling from Raw Audio
2021
Top co-authors
Yossi Adi
· 16
Gabriel Synnaeve
· 9
Felix Kreuk
· 7
Abdelrahman Mohamed
· 6
Adam Polyak
· 5
Alexandre D\'efossez
· 5
Itai Gat
· 5
Eugene Kharitonov
· 4
Morgane Rivière
· 4
Tal Remez
· 4
Robin Algayres
· 3
Ann Lee
· 2
Topics
Speech Recognition
Audio Generation
Audio Understanding
Text-to-Speech
Multimodal Audio
Speech Enhancement
Speaker Analysis
Music Generation
Speech Translation
Voice Cloning