← all papers · overview

Lapdoc: Layout-aware Prompting For Documents

Abstract

Recent advances in training large language models (LLMs) using massive amounts of solely textual data lead to strong generalization across many domains and tasks, including document-specific tasks. Opposed to that there is a trend to train multi-modal transformer architectures tailored for document understanding that are designed specifically to fuse textual inputs with the corresponding document

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).