Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Po-chun Hsu — most-cited papers & profile · Speech Audio
← authors
·
overview
Po-chun Hsu
16
papers ·
58
citations
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Rhythm-Flexible Voice Conversion without Parallel Data Using Cycle-GAN over Phoneme Posteriorgram Sequences
2018 · 25 citations
Towards Robust Neural Vocoding for Speech Generation: A Survey
2019 · 23 citations
Silence is Sweeter Than Speech: Self-Supervised Model Using Silence to Store Speaker Information
2022 · 6 citations
WG-WaveNet: Real-Time High-Fidelity Speech Synthesis without GPU
2020 · 2 citations
Investigating on Incorporating Pretrained and Learnable Speaker Representations for Multi-Speaker Multi-Style Text-to-Speech
2021 · 1 citations
Low-Resource Self-Supervised Learning with SSL-Enhanced TTS
2023 · 1 citations
Unsupervised End-to-End Learning of Discrete Linguistic Units for Voice Conversion
2019
Mockingjay: Unsupervised Speech Representation Learning with Deep Bidirectional Transformer Encoders
2019
Universal Adaptor: Converting Mel-Spectrograms Between Different Configurations for Speech Synthesis
2022
Parallel Synthesis for Autoregressive Speech Generation
2022
STOP: A dataset for Spoken Task Oriented Semantic Parsing
2022
Learning Phone Recognition from Unpaired Audio and Phone Sequences Based on Generative Adversarial Network
2022
Let's Fuse Step by Step: A Generative Fusion Decoding Algorithm with LLMs for Robust and Instruction-Aware ASR and OCR
2024
Building a Taiwanese Mandarin Spoken Language Model: A First Attempt
2024
Top co-authors
Hung-yi Lee
· 7
Da-Rong Liu
· 3
Hung-yi Lee
· 3
Abdelrahman Mohamed
· 2
Ali Elkahky
· 2
and Hung-yi Lee
· 2
Andy T. Liu
· 2
Andy T. Liu
· 2
Chan-Jan Hsu
· 2
Emmanuel Dupoux
· 2
Jade Copet
· 2
Shu-wen Yang
· 2
Topics
Speech Recognition
Text-to-Speech
Audio Generation
Voice Cloning
Audio Understanding
Speaker Analysis
Speech Enhancement
Music Generation
Speech Translation
Multimodal Audio