Papers
Topics
Authors
Recent
Gemini 2.5 Flash
Gemini 2.5 Flash
110 tokens/sec
GPT-4o
56 tokens/sec
Gemini 2.5 Pro Pro
44 tokens/sec
o3 Pro
6 tokens/sec
GPT-4.1 Pro
47 tokens/sec
DeepSeek R1 via Azure Pro
28 tokens/sec
2000 character limit reached

StorySparkQA: Expert-Annotated QA Pairs with Real-World Knowledge for Children's Story-Based Learning (2311.09756v3)

Published 16 Nov 2023 in cs.CL

Abstract: Interactive story reading is a common parent-child activity, where parents expect to teach both language skills and real-world knowledge beyond the story. While increasing storytelling and reading systems have been developed for this activity, they often fail to infuse real-world knowledge into the conversation. This limitation can be attributed to the existing question-answering (QA) datasets used for children's education, upon which the systems are built, failing to capture the nuances of how education experts think when conducting interactive story reading activities. To bridge this gap, we design an annotation framework, empowered by existing knowledge graph to capture experts' annotations and thinking process, and leverage this framework to construct StorySparkQA dataset, which comprises 5,868 expert-annotated QA pairs with real-world knowledge. We conduct automated and human expert evaluations across various QA pair generation settings to demonstrate that our StorySparkQA can effectively support models in generating QA pairs that target real-world knowledge beyond story content. StorySparkQA is available at https://huggingface.co/datasets/NEU-HAI/StorySparkQA.

User Edit Pencil Streamline Icon: https://streamlinehq.com
Authors (10)
  1. Jiaju Chen (10 papers)
  2. Yuxuan Lu (26 papers)
  3. Shao Zhang (18 papers)
  4. Bingsheng Yao (49 papers)
  5. Yuanzhe Dong (4 papers)
  6. Ying Xu (81 papers)
  7. Yunyao Li (43 papers)
  8. Qianwen Wang (17 papers)
  9. Dakuo Wang (87 papers)
  10. Yuling Sun (10 papers)
Citations (3)

Summary

We haven't generated a summary for this paper yet.