Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Zhiyao Duan — most-cited papers & profile · Speech Audio
← authors
·
overview
Zhiyao Duan
24
papers ·
66
citations ·
28
h-index
University of Rochester
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
SVDD 2024: The Inaugural Singing Voice Deepfake Detection Challenge
2024 · 11 citations
UR Channel-Robust Synthetic Speech Detection System for ASVspoof 2021
2021 · 4 citations
SingNet: A Real-time Singing Voice Beat and Downbeat Tracking System
2023 · 3 citations
SVDD Challenge 2024: A Singing Voice Deepfake Detection Challenge Evaluation Plan
2024 · 3 citations
ControlVC: Zero-Shot Voice Conversion with Time-Varying Controls on Pitch and Speed
2022 · 1 citations
Y-Vector: Multiscale Waveform Encoder for Speaker Embedding
2020 · 1 citations
Learning Sparse Analytic Filters for Piano Transcription
2021 · 1 citations
Singing Beat Tracking With Self-supervised Front-end and Linear Transformers
2022 · 1 citations
Generating Novel and Realistic Speakers for Voice Conversion
2025
Conan: A Chunkwise Online Network for Zero-Shot Adaptive Voice Conversion
2025
PartialEdit: Identifying Partial Deepfakes in the Era of Neural Speech Editing
2025
Spoofing Speaker Verification Systems with Deep Multi-speaker Text-to-speech Synthesis
2019
RL-Duet: Online Music Accompaniment Generation Using Deep Reinforcement Learning
2020
A study of the robustness of raw waveform based speaker embeddings under mismatched conditions
2021
SAMO: Speaker Attractor Multi-Center One-Class Learning for Voice Anti-Spoofing
2022
Top co-authors
You Zhang
· 5
Ge Zhu
· 4
Frank Cwitkowitz
· 3
Jiatong Shi
· 3
Tomoki Toda
· 3
Yongyi Zang
· 3
Baotong Tian
· 2
Ge Zhu
· 2
Jionghao Han
· 2
Juan-Pablo Cáceres
· 2
Meiying Melissa Chen
· 2
Mojtaba Heydari
· 2
Topics
Audio Understanding
Audio Generation
Music Generation
Voice Cloning
Speech Recognition
Text-to-Speech
Speaker Analysis
Speech Enhancement
Speech Translation