Papers
Topics
Authors
Recent
Gemini 2.5 Flash
Gemini 2.5 Flash
110 tokens/sec
GPT-4o
56 tokens/sec
Gemini 2.5 Pro Pro
44 tokens/sec
o3 Pro
6 tokens/sec
GPT-4.1 Pro
47 tokens/sec
DeepSeek R1 via Azure Pro
28 tokens/sec
2000 character limit reached

KPDrop: Improving Absent Keyphrase Generation (2112.01476v3)

Published 2 Dec 2021 in cs.CL

Abstract: Keyphrase generation is the task of generating phrases (keyphrases) that summarize the main topics of a given document. Keyphrases can be either present or absent from the given document. While the extraction of present keyphrases has received much attention in the past, only recently a stronger focus has been placed on the generation of absent keyphrases. However, generating absent keyphrases is challenging; even the best methods show only a modest degree of success. In this paper, we propose a model-agnostic approach called keyphrase dropout (or KPDrop) to improve absent keyphrase generation. In this approach, we randomly drop present keyphrases from the document and turn them into artificial absent keyphrases during training. We test our approach extensively and show that it consistently improves the absent performance of strong baselines in both supervised and resource-constrained semi-supervised settings.

User Edit Pencil Streamline Icon: https://streamlinehq.com
Authors (4)
  1. Jishnu Ray Chowdhury (17 papers)
  2. Seoyeon Park (2 papers)
  3. Tuhin Kundu (3 papers)
  4. Cornelia Caragea (58 papers)
Citations (6)

Summary

We haven't generated a summary for this paper yet.