Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Tom Ko — most-cited papers & profile · Speech Audio
← authors
·
overview
Tom Ko
22
papers ·
130
citations ·
17
h-index
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
SpeechT5: Unified-Modal Encoder-Decoder Pre-Training for Spoken Language Processing
2021 · 30 citations
M3ST: Mix at Three Levels for Speech Translation
2022 · 14 citations
Pre-Training Transformer Decoder for End-to-End ASR Model with Unpaired Speech Data
2022 · 13 citations
An Investigation of Few-Shot Learning in Spoken Term Classification
2018 · 9 citations
PolyVoice: Language Models for Speech to Speech Translation
2023 · 5 citations
Exploring Machine Speech Chain for Domain Adaptation and Few-Shot Speaker Adaptation
2021 · 4 citations
Speech Translation with Large Language Models: An Industrial Practice
2023 · 4 citations
Autospeech 2020: The Second Automated Machine Learning Challenge For Speech Classification
2020 · 3 citations
Multi-View Self-Attention Based Transformer for Speaker Recognition
2021 · 3 citations
GigaST: A 10,000-hour Pseudo Speech Translation Corpus
2022 · 2 citations
Leveraging Pseudo-labeled Data to Improve Direct Speech-to-Speech Translation
2022 · 2 citations
DUB: Discrete Unit Back-translation for Speech Translation
2023 · 2 citations
Recent Advances in Direct Speech-to-text Translation
2023 · 2 citations
AutoSpeech 2020: The Second Automated Machine Learning Challenge for Speech Classification
2020 · 1 citations
Towards Achieving Human Parity on End-to-end Simultaneous Speech Translation via LLM Agent
2024 · 1 citations
Top co-authors
Mingxuan Wang
· 9
Haizhou Li
· 4
Long Zhou
· 4
Rong Ye
· 4
Yu Zhang
· 4
Chen Xu
· 3
Qing Li
· 3
Rui Wang
· 3
Shujie Liu
· 3
Zhichao Huang
· 3
Zhihua Wei
· 3
Furu Wei
· 2
Topics
Speech Recognition
Text-to-Speech
Speech Translation
Audio Understanding
Speech Enhancement
Speaker Analysis
eess.AS
cs.AI
cs.SD
Voice Cloning