← all papers · overview

Tackling VQA With Pretrained Foundation Models Without Further Training

Abstract

Large language models (LLMs) have achieved state-of-the-art results in many natural language processing tasks. They have also demonstrated ability to adapt well to different tasks through zero-shot or few-shot settings. With the capability of these LLMs, researchers have looked into how to adopt them for use with Visual Question Answering (VQA). Many methods require further training to align the i

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).