Papers
Topics
Authors
Recent
Gemini 2.5 Flash
Gemini 2.5 Flash
110 tokens/sec
GPT-4o
56 tokens/sec
Gemini 2.5 Pro Pro
44 tokens/sec
o3 Pro
6 tokens/sec
GPT-4.1 Pro
47 tokens/sec
DeepSeek R1 via Azure Pro
28 tokens/sec
2000 character limit reached

On Code Reuse from StackOverflow: An Exploratory Study on Jupyter Notebook (2302.11732v1)

Published 23 Feb 2023 in cs.SE

Abstract: Jupyter Notebook is a popular tool among data analysts and scientists for working with data. It provides a way to combine code, documentation, and visualizations in a single, interactive environment, facilitating code reuse. While code reuse can improve programming efficiency, it can also decrease readability, security, and overall performance. We conduct a large-scale exploratory study of code reuse practices in the Jupyter Notebook development community on the Stack Overflow platform to understand the potential negative impacts of code reuse. Our findings identified 1,097,470 Jupyter Notebook clone pairs that reuse Stack Overflow code snippets, and the average code snippet has 7.91 code quality violations. Through our research, we gain insight into the reasons behind Jupyter Notebook developers' decision to reuse code and the potential drawbacks of this practice.

User Edit Pencil Streamline Icon: https://streamlinehq.com
Authors (4)
  1. Mingke Yang (2 papers)
  2. Yuming Zhou (19 papers)
  3. Bixin Li (3 papers)
  4. Yutian Tang (17 papers)
Citations (2)

Summary

We haven't generated a summary for this paper yet.