Awesome Generative Models
π
Papers
π§
Topics
π₯
Trending
πΊοΈ
Map
π
Leaderboards
π
Learn
π€
Ask AI
β―
More
π₯
Authors
π
Reading Packs
π
Datasets
π οΈ
Tools
π°
News
π
Blogs
βοΈ
Newsletter
π―
Research Radar
π
Saved
+ Add Paper
βΎ
β
β authors
Β·
overview
Loading authorβ¦
π€
Ask AI
Shao-Yen Tseng β most-cited papers & profile Β· Generative Models
β authors
Β·
overview
Shao-Yen Tseng
7
papers Β·
8
citations Β·
10
h-index
Google Scholar β
Semantic Scholar β
OpenAlex β
Most-cited papers
KD-VLP: Improving End-to-End Vision-and-Language Pretraining with Object Knowledge Distillation
2021 Β· 4 citations
VL-InterpreT: An Interactive Visualization Tool for Interpreting Vision-Language Transformers
2022 Β· 3 citations
LLaVA-Gemma: Accelerating Multimodal Foundation Models with a Compact Language Model
2024 Β· 1 citations
LieCraft: A Multi-Agent Framework for Evaluating Deceptive Capabilities in Language Models
2026
MuMUR : Multilingual Multimodal Universal Retrieval
2022
ManagerTower: Aggregating the Insights of Uni-Modal Experts for Vision-Language Representation Learning
2023
FiVL: A Framework for Improved Vision-Language Alignment through the Lens of Training, Evaluation and Explainability
2024
Topics
Vision-Language Models
Video-Language
Evaluation
Image-Text Retrieval
Visual QA & Reasoning
Multi-Agent
Safety
cs.MM
Vision-Language
Model Architecture