TVQA
Emerging10papers using it
2019first seen
TVQA is a dataset that contains video clips paired with questions and answers, used to evaluate the performance of models in video question answering tasks.
Papers using TVQA (10)
- Multi-speaker Attention Alignment for Multimodal Social InteractionPOVQA: Preference-optimized Video Question Answering With Rationales For Data EfficiencyMultimodal Conversation Structure UnderstandingZero-Shot Video Question Answering via Frozen Bidirectional Language
ModelsVideo Question Generation via Cross-Modal Self-Attention Networks
LearningModality Shifting Attention Network for Multi-modal Video Question
AnsweringFrame-Subtitle Self-Supervision for Multi-Modal Video Question AnsweringMERLOT Reserve: Neural Script Knowledge through Vision and Language and
SoundTV-TREES: Multimodal Entailment Trees for Neuro-Symbolic Video ReasoningOn Modality Bias in the TVQA Dataset