Papers
Topics
Authors
Recent
Gemini 2.5 Flash
Gemini 2.5 Flash
119 tokens/sec
GPT-4o
56 tokens/sec
Gemini 2.5 Pro Pro
43 tokens/sec
o3 Pro
6 tokens/sec
GPT-4.1 Pro
47 tokens/sec
DeepSeek R1 via Azure Pro
28 tokens/sec
2000 character limit reached

Using Inter-Sentence Diverse Beam Search to Reduce Redundancy in Visual Storytelling (1805.11867v1)

Published 30 May 2018 in cs.CL and cs.AI

Abstract: Visual storytelling includes two important parts: coherence between the story and images as well as the story structure. For image to text neural network models, similar images in the sequence would provide close information for story generator to obtain almost identical sentence. However, repeatedly narrating same objects or events will undermine a good story structure. In this paper, we proposed an inter-sentence diverse beam search to generate a more expressive story. Comparing to some recent models of visual storytelling task, which generate story without considering the generated sentence of the previous picture, our proposed method can avoid generating identical sentence even given a sequence of similar pictures.

User Edit Pencil Streamline Icon: https://streamlinehq.com
Authors (4)
  1. Chao-Chun Hsu (13 papers)
  2. Szu-Min Chen (1 paper)
  3. Ming-Hsun Hsieh (1 paper)
  4. Lun-Wei Ku (35 papers)
Citations (16)

Summary

We haven't generated a summary for this paper yet.