Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Hongyu Gong — most-cited papers & profile · Speech Audio
← authors
·
overview
Hongyu Gong
18
papers ·
79
citations ·
34
h-index
Shandong University · Hong Kong University of Science and Technology · Inner Mongolia University · Ministry of Education · University of Hong Kong
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Seamless: Multilingual Expressive and Streaming Speech Translation
2023 · 41 citations
SeamlessM4T: Massively Multilingual & Multimodal Machine Translation
2023 · 13 citations
Pre-training for Speech Translation: CTC Meets Optimal Transport
2023 · 7 citations
Speech-to-Speech Translation For A Real-world Unwritten Language
2022 · 5 citations
SpeechMatrix: A Large-Scale Mined Corpus of Multilingual Speech-to-Speech Translations
2022 · 4 citations
Multilingual Speech-to-Speech Translation into Multiple Target Languages
2023 · 4 citations
Direct Simultaneous Speech-to-Speech Translation with Variational Monotonic Multihead Attention
2021 · 1 citations
Textless Speech-to-Speech Translation on Real Data
2021 · 1 citations
A Holistic Cascade System, benchmark, and Human Evaluation Protocol for Expressive Speech-to-Speech Translation
2023 · 1 citations
Exploration on HuBERT with Multiple Resolutions
2023 · 1 citations
MSLM-S2ST: A Multitask Speech Language Model for Textless Speech-to-Speech Translation with Speaker Style Preservation
2024 · 1 citations
From Start to Finish: Latency Reduction Strategies for Incremental Speech Synthesis in Simultaneous Speech-to-Speech Translation
2021
Unified Speech-Text Pre-training for Speech Translation and Recognition
2022
Improving Speech-to-Speech Translation Through Unlabeled Text
2022
An Empirical Study of Speech Language Models for Prompt-Conditioned Speech Synthesis
2024
Top co-authors
Changhan Wang
· 13
Juan Pino
· 11
Ann Lee
· 8
Ilia Kulikov
· 7
Peng-Jen Chen
· 7
Hirofumi Inaguma
· 5
Holger Schwenk
· 5
Ning Dong
· 5
Paul-Ambroise Duquenne
· 5
Sravya Popuri
· 5
Yilin Yang
· 5
Yun Tang
· 5
Topics
Text-to-Speech
Speech Recognition
Speech Translation
Audio Understanding
Speaker Analysis
Multimodal Audio
Audio Generation
Speech Enhancement