Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Neil Houlsby — most-cited papers & profile · Multimodal
← authors
·
overview
Neil Houlsby
20
papers ·
23529
citations ·
38
h-index
Google (United States)
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
2020 · 21717 citations
MLP-Mixer: An all-MLP Architecture for Vision
2021 · 1445 citations
Scaling Vision Transformers to 22 Billion Parameters
2023 · 118 citations
Simple Open-Vocabulary Object Detection with Vision Transformers
2022 · 64 citations
On Self Modulation for Generative Adversarial Networks
2018 · 44 citations
PaLI-X: On Scaling up a Multilingual Vision and Language Model
2023 · 39 citations
Scaling Vision with Sparse Mixture of Experts
2021 · 29 citations
Learning to Merge Tokens in Vision Transformers
2022 · 24 citations
Scaling Vision Transformers
2021 · 16 citations
Patch n' Pack: NaViT, a Vision Transformer for any Aspect Ratio and Resolution
2023 · 16 citations
Image Captioners Are Scalable Vision Learners Too
2023 · 10 citations
Self-Supervised GANs via Auxiliary Rotation Loss
2018 · 7 citations
Experimental Adaptive Bayesian Tomography
2013
Self-Supervised Learning of Video-Induced Visual Invariances
2019
Representation learning from videos in-the-wild: An object-centric approach
2020
Topics
3D Vision
Video Understanding
Object Detection
Visual Language
GANs
Conditioning & Control
Image Restoration
Image Generation
Segmentation
Model Architecture