Awesome Large Language Models
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Marcella Cornia — most-cited papers & profile · Large Language Models
← authors
·
overview
Marcella Cornia
20
papers ·
123
citations ·
17
h-index
University of Modena and Reggio Emilia
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Retrieval-Augmented Transformer for Image Captioning
2022 · 48 citations
CaMEL: Mean Teacher Learning for Image Captioning
2022 · 38 citations
Wiki-LLaVA: Hierarchical Retrieval-Augmented Generation for Multimodal LLMs
2024 · 35 citations
Learning to Mask and Permute Visual Tokens for Vision Transformer Pre-Training
2023 · 2 citations
Segmenting, Fast and Slow: Real-Time Open-Vocabulary Video Instance Segmentation with Dual-Path Processing
2026
Mind the Heads: Topological Representation Alignment for Multimodal LLMs
2026
Recurrence Meets Transformers For Universal Multimodal Retrieval
2025
Look Twice: Training-Free Evidence Highlighting in Multimodal Large Language Models
2026
CounterVid: Counterfactual Video Generation for Mitigating Action and Temporal Hallucinations in Video-Language Models
2026
ReAG: Reasoning-Augmented Generation for Knowledge-based Visual Question Answering
2025
Multimodal Attention Networks for Low-Level Vision-and-Language Navigation
2019
A Novel Attention-based Aggregation Function to Combine Vision and Language
2020
ALADIN: Distilling Fine-grained Alignment Scores for Efficient Image-Text Matching and Retrieval
2022
ALADIN: Distilling Fine-grained Alignment Scores for Efficient Image-Text Matching and Retrieval
2022
Positive-Augmented Contrastive Learning for Image and Video Captioning Evaluation
2023
Topics
Vision-Language Models
Benchmarks
Visual Language
Visual QA & Reasoning
Video-Language
Image Generation
Image-Text Retrieval
Video Understanding
Audio-Visual
3D Vision