Awesome Computer Vision
Papers
Topics
Trending
Leaderboards
Authors
Datasets
Learn
Ask AI
Map
Tools
Reading Packs
News
All sections
Browse
Papers
The full index, filterable and sortable.
Topics
The same papers grouped by subject tag.
Trending
What moved this week, and by how much.
Authors
Who publishes here, and who they publish with.
Map
The collection laid out by embedding similarity.
Compare
Leaderboards
Benchmark tables, with the paper behind each number.
Datasets
The datasets these papers train and evaluate on.
Tools
Code and libraries released alongside the papers.
Read
Learn
Ordered routes from background reading to current work.
Ask AI
Ask a question and get answers cited to these papers.
Videos
The most-watched talks and lectures in this field.
Reading Packs
Short curated sets built around one question.
Follow
News
Press and coverage tied back to the papers.
Blogs
Author and lab write-ups of their own work.
Newsletter
Email digest of what changed, on a schedule.
Research Radar
Paste an abstract, get matches across every collection.
Yours
Saved
Papers you bookmarked in this browser.
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
Ask AI
Aman Chadha — most-cited papers & profile · Computer Vision
← authors
·
overview
Aman Chadha
33
papers ·
78
citations ·
0
h-index
Amazon (United States) · San Diego State University · Amazon (Germany) · Apple (United Kingdom) · Apple (United States) · Stanford University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
iPerceive: Applying Common-Sense Reasoning to Multi-Modal Dense Video Captioning and Video Question Answering
2020 · 24 citations
iReason: Multimodal Commonsense Reasoning using Videos and Natural Language with Interpretability
2021 · 4 citations
Seeing the Pose in the Pixels: Learning Pose-Aware Representations in Vision Transformers
2023 · 4 citations
How Culturally Aware are Vision-Language Models?
2024 · 4 citations
Density Adaptive Attention is All You Need: Robust Parameter-Efficient Fine-Tuning Across Multiple Modalities
2024 · 2 citations
Facial Expression Recognition using Squeeze and Excitation-powered Swin Transformers
2023 · 1 citations
The Visual Counter Turing Test (VCT2): A Benchmark for Evaluating AI-Generated Image Detection and the Visual AI Index (VAI)
2024 · 1 citations
I see what you hear: a vision-inspired method to localize words
2022
Few-shot Multimodal Multitask Multilingual Learning
2023
IMAGINATOR: Pre-Trained Image+Text Joint Embeddings using Word-Level Grounding of Images
2023
Refining Text-to-Image Generation: Towards Accurate Training-Free Glyph-Enhanced Image Generation
2024
MedVisionLlama: Leveraging Pre-Trained Large Language Model Layers to Enhance Medical Image Segmentation
2024
Top co-authors
Vinija Jain
· 5
Amitava Das
· 2
Amit Sheth
· 2
Aaron Elkins
· 1
Abhilekh Borah
· 1
Amir Shmuel
· 1
Arnav Kundu
· 1
Arpita Vats
· 1
Ashhar Aziz
· 1
Ashish Shrivastava
· 1
Devang Naik
· 1
Dominick Reilly
· 1
Topics
cs.CV
cs.AI
cs.LG
cs.CL
cs.MM
eess.IV
cs.SD
eess.AS
eess.SP
Image Generation