Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Xiaodong He — most-cited papers & profile · Multimodal
← authors
·
overview
Xiaodong He
40
papers ·
2551
citations ·
71
h-index
Northeast Agricultural University · Shandong Academy of Agricultural Machinery Sciences · Jingdong (China)
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
MS-Celeb-1M: A Dataset and Benchmark for Large-Scale Face Recognition
2016 · 1998 citations
Adversarial Ranking for Language Generation
2017 · 158 citations
Bottom-Up and Top-Down Attention for Image Captioning and Visual Question Answering
2017 · 94 citations
End-to-end Structure-Aware Convolutional Networks for Knowledge Base Completion
2018 · 62 citations
Multiple instance learning with graph neural networks
2019 · 55 citations
Object-driven Text-to-Image Synthesis via Adversarial Training
2019 · 46 citations
Rich Image Captioning in the Wild
2016 · 41 citations
On the Discrimination-Generalization Tradeoff in GANs
2017 · 34 citations
Hierarchically Structured Reinforcement Learning for Topically Coherent Visual Story Generation
2018 · 21 citations
Towards adversarial learning of speaker-invariant representation for speech emotion recognition
2019 · 17 citations
Deep Reinforcement Learning with a Combinatorial Action Space for Predicting Popular Reddit Threads
2016 · 9 citations
Constrained Convolutional-recurrent Networks To Improve Speech Quality With Low Impact On Recognition Accuracy
2018 · 6 citations
Reinforcement Learning To Adapt Speech Enhancement to Instantaneous Input Signal Quality
2017 · 2 citations
Reinforcement Learning To Adapt Speech Enhancement to Instantaneous Input Signal Quality
2017 · 2 citations
Deep Speaker Embedding Learning with Multi-Level Pooling for Text-Independent Speaker Verification
2019 · 2 citations
Topics
Speech Recognition
Control
Model-Based RL
Hardware
Speech Translation
Speech Enhancement
Policy Gradient
Text-to-Speech
Human-Robot Interaction
GANs