Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Daxin Tan — most-cited papers & profile · Speech Audio
← authors
·
overview
Daxin Tan
15
papers ·
8
citations ·
5
h-index
Artificial Intelligence in Medicine (Canada)
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
CUHK-EE Voice Cloning System for ICASSP 2021 M2VoC Challenge
2021 · 4 citations
EditSpeech: A Text Based Speech Editing System Using Partial Inference and Bidirectional Fusion
2021 · 1 citations
Environment Aware Text-to-Speech Synthesis
2021 · 1 citations
Mixed-Phoneme BERT: Improving BERT with Mixed Phoneme and Sup-Phoneme Representations for Text to Speech
2022 · 1 citations
Enhancing Code-switched Text-to-Speech Synthesis Capability in Large Language Models with only Monolingual Corpora
2024 · 1 citations
SpeechEditBench: A Bilingual Multi-Attribute Benchmark for Instruction-Guided Speech Editing
2026
Minimizing Modality Gap from the Input Side: Your Speech LLM Can Be a Prosody-Aware Text LLM
2026
Speech-Omni-Lite: Portable Speech Interfaces for Vision-Language Models
2026
DSA-Tokenizer: Disentangled Semantic-Acoustic Tokenization via Flow Matching-based Hierarchical Fusion
2026
PROST-LLM: Progressively Enhancing the Speech-to-Speech Translation Capability in LLMs
2026
EMOVA: Empowering Language Models to See, Hear and Speak with Vivid Emotions
2024
Applying the Information Bottleneck Principle to Prosodic Representation Learning
2021
A study on the efficacy of model pre-training in developing neural text-to-speech system
2021
ToneUnit: A Speech Discretization Approach for Tonal Language Speech Synthesis
2024
Exploring SSL Discrete Tokens for Multilingual ASR
2024
Top co-authors
Tan Lee
· 7
Guangyan Zhang
· 5
Dingdong Wang
· 2
Kaitao Song
· 2
Sheng Zhao
· 2
Xiao Chen
· 2
Xu Tan
· 2
Yu Ting Yeung
· 2
Chunwei Wang
· 1
Dehua Tao
· 1
Dehua Tao
· 1
Dehua Tao
· 1
Topics
Text-to-Speech
Speech Recognition
Audio Generation
Audio Understanding
Multimodal Audio
Speech Enhancement
Voice Cloning
Speech Translation
Speaker Analysis