2000 character limit reached
Recent Advances in Video Question Answering: A Review of Datasets and Methods (2101.05954v2)
Published 15 Jan 2021 in cs.CV
Abstract: Video Question Answering (VQA) is a recent emerging challenging task in the field of Computer Vision. Several visual information retrieval techniques like Video Captioning/Description and Video-guided Machine Translation have preceded the task of VQA. VQA helps to retrieve temporal and spatial information from the video scenes and interpret it. In this survey, we review a number of methods and datasets for the task of VQA. To the best of our knowledge, no previous survey has been conducted for the VQA task.
- Devshree Patel (2 papers)
- Ratnam Parikh (2 papers)
- Yesha Shastri (2 papers)