Papers
Topics
Authors
Recent
Gemini 2.5 Flash
Gemini 2.5 Flash
102 tokens/sec
GPT-4o
59 tokens/sec
Gemini 2.5 Pro Pro
43 tokens/sec
o3 Pro
6 tokens/sec
GPT-4.1 Pro
50 tokens/sec
DeepSeek R1 via Azure Pro
28 tokens/sec
2000 character limit reached

Transferring Domain-Agnostic Knowledge in Video Question Answering (2110.13395v1)

Published 26 Oct 2021 in cs.CV and cs.AI

Abstract: Video question answering (VideoQA) is designed to answer a given question based on a relevant video clip. The current available large-scale datasets have made it possible to formulate VideoQA as the joint understanding of visual and language information. However, this training procedure is costly and still less competent with human performance. In this paper, we investigate a transfer learning method by the introduction of domain-agnostic knowledge and domain-specific knowledge. First, we develop a novel transfer learning framework, which finetunes the pre-trained model by applying domain-agnostic knowledge as the medium. Second, we construct a new VideoQA dataset with 21,412 human-generated question-answer samples for comparable transfer of knowledge. Our experiments show that: (i) domain-agnostic knowledge is transferable and (ii) our proposed transfer learning framework can boost VideoQA performance effectively.

User Edit Pencil Streamline Icon: https://streamlinehq.com
Authors (6)
  1. Tianran Wu (2 papers)
  2. Noa Garcia (33 papers)
  3. Mayu Otani (32 papers)
  4. Chenhui Chu (48 papers)
  5. Yuta Nakashima (67 papers)
  6. Haruo Takemura (8 papers)
Citations (7)

Summary

We haven't generated a summary for this paper yet.