Unprecedented Code Change Automation: The Fusion of LLMs and Transformation by Example (2402.07138v3)

Published 11 Feb 2024 in cs.SE

Abstract: Software developers often repeat code changes, known as "code change patterns" (CPATs), within and across projects. Automating these CPATs accelerates development, but current Transformation by Example (TBE) techniques are limited by the input examples' quality and quantity, missing variations with different syntax or flow yet semantically similar. LLMs, trained on vast code datasets, can overcome these limitations by generating semantically equivalent, unseen CPAT variants, enhancing TBE effectiveness. We identified best practices for using LLMs to generate code variants meeting criteria of correctness, usefulness, and applicability. Implementing these in PyCraft, combining static and dynamic analysis with LLMs, we achieved an F-measure of 96.6% in identifying correct variants, expanding inputs by 58x on average, and automating changes to increase target codes by up to 39x. Patches from PyCraft were submitted to projects like microsoft/DeepSpeed and IBM/inFairness, with an 83% acceptance rate, validating our approach's usefulness.

References (66)

Citations (10)

View on Semantic Scholar

Summary

We haven't generated a summary for this paper yet.

Summarize Now

Follow-up Questions

We haven't generated follow-up questions for this paper yet.

Generate Now

Authors (4)

Tweets

https://twitter.com/areyde/status/1813570300292088263

https://twitter.com/ComputerPapers/status/1757316232607342874

https://twitter.com/ComputerPapers/status/1787913676684468286

Unprecedented Code Change Automation: The Fusion of LLMs and Transformation by Example (2402.07138v3)

Summary

Follow-up Questions

Related Papers

Authors (4)

Tweets