Papers
Topics
Authors
Recent
Gemini 2.5 Flash
Gemini 2.5 Flash
119 tokens/sec
GPT-4o
56 tokens/sec
Gemini 2.5 Pro Pro
43 tokens/sec
o3 Pro
6 tokens/sec
GPT-4.1 Pro
47 tokens/sec
DeepSeek R1 via Azure Pro
28 tokens/sec
2000 character limit reached

QuTI! Quantifying Text-Image Consistency in Multimodal Documents (2104.13748v1)

Published 28 Apr 2021 in cs.IR and cs.MM

Abstract: The World Wide Web and social media platforms have become popular sources for news and information. Typically, multimodal information, e.g., image and text is used to convey information more effectively and to attract attention. While in most cases image content is decorative or depicts additional information, it has also been leveraged to spread misinformation and rumors in recent years. In this paper, we present a Web-based demo application that automatically quantifies the cross-modal relations of entities (persons, locations, and events) in image and text. The applications are manifold. For example, the system can help users to explore multimodal articles more efficiently, or can assist human assessors and fact-checking efforts in the verification of the credibility of news stories, tweets, or other multimodal documents.

User Edit Pencil Streamline Icon: https://streamlinehq.com
Authors (3)
  1. Matthias Springstein (7 papers)
  2. Eric Müller-Budack (19 papers)
  3. Ralph Ewerth (61 papers)
Citations (10)

Summary

We haven't generated a summary for this paper yet.