Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Hexiang Hu — most-cited papers & profile · Multimodal
← authors
·
overview
Hexiang Hu
13
papers ·
178
citations ·
27
h-index
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Multimodal Model-Agnostic Meta-Learning via Task-Aware Modulation
2019 · 72 citations
PaLI-X: On Scaling up a Multilingual Vision and Language Model
2023 · 39 citations
On Model Calibration for Long-Tailed Object Detection and Instance Segmentation
2021 · 27 citations
LabelBank: Revisiting Global Perspectives for Semantic Segmentation
2017 · 18 citations
Visual Storytelling via Predicting Anchor Word Embeddings in the Stories
2020 · 5 citations
PreSTU: Pre-Training for Scene-Text Understanding
2022 · 5 citations
Synthesized Policies for Transfer and Adaptation across Tasks and Environments
2019 · 4 citations
Imagen 3
2024 · 4 citations
Recalling Holistic Information for Semantic Segmentation
2016 · 2 citations
FastMask: Segment Multi-scale Object Candidates in One Shot
2016 · 2 citations
Inference-Time Scaling for Diffusion Models beyond Scaling Denoising Steps
2025
Instruct-Imagen: Image Generation with Multi-modal Instruction
2024
MEGA-Bench: Scaling Multimodal Evaluation to over 500 Real-World Tasks
2024
Topics
Segmentation
Object Detection
Visual Language
Training Techniques
Meta-RL
Video Understanding
Vision-Language
Evaluation
Efficiency
Model Architecture