Papers
Topics
Authors
Recent
Gemini 2.5 Flash
Gemini 2.5 Flash
41 tokens/sec
GPT-4o
59 tokens/sec
Gemini 2.5 Pro Pro
41 tokens/sec
o3 Pro
7 tokens/sec
GPT-4.1 Pro
50 tokens/sec
DeepSeek R1 via Azure Pro
28 tokens/sec
2000 character limit reached

Weibo-COV: A Large-Scale COVID-19 Social Media Dataset from Weibo (2005.09174v6)

Published 19 May 2020 in cs.SI

Abstract: With the rapid development of COVID-19 around the world, people are requested to maintain "social distance" and "stay at home". In this scenario, extensive social interactions transfer to cyberspace, especially on social media platforms like Twitter and Sina Weibo. People generate posts to share information, express opinions and seek help during the pandemic outbreak, and these kinds of data on social media are valuable for studies to prevent COVID-19 transmissions, such as early warning and outbreaks detection. Therefore, in this paper, we release a novel and fine-grained large-scale COVID-19 social media dataset collected from Sina Weibo, named Weibo-COV, contains more than 40 million posts ranging from December 1, 2019 to April 30, 2020. Moreover, this dataset includes comprehensive information nuggets like post-level information, interactive information, location information, and repost network. We hope this dataset can promote studies of COVID-19 from multiple perspectives and enable better and rapid researches to suppress the spread of this pandemic.

User Edit Pencil Streamline Icon: https://streamlinehq.com
Authors (4)
  1. Yong Hu (116 papers)
  2. Heyan Huang (107 papers)
  3. Anfan Chen (5 papers)
  4. Xian-Ling Mao (76 papers)
Citations (49)