Papers

Topics

Authors

Recent

View all

Detailed Answer

Quick Answer

Concise responses based on abstracts only

Detailed Answer

Well-researched responses based on abstracts and relevant paper content.

Custom Instructions Pro

Preferences or requirements that you'd like Emergent Mind to consider when generating responses

Gemini 2.5 Flash

Gemini 2.5 Flash 87 tok/s

Gemini 2.5 Pro 47 tok/s Pro

GPT-5 Medium 29 tok/s Pro

GPT-5 High 37 tok/s Pro

GPT-4o 85 tok/s Pro

Kimi K2 183 tok/s Pro

GPT OSS 120B 419 tok/s Pro

Claude Sonnet 4 37 tok/s Pro

2000 character limit reached

Semantic-Preserving Transformations as Mutation Operators: A Study on Their Effectiveness in Defect Detection (2503.23448v1)

Published 30 Mar 2025 in cs.SE, cs.AI, cs.CL, and cs.LG

Abstract: Recent advances in defect detection use LLMs. Existing works enhanced the training data to improve the models' robustness when applied to semantically identical code (i.e., predictions should be the same). However, the use of semantically identical code has not been considered for improving the tools during their application - a concept closely related to metamorphic testing. The goal of our study is to determine whether we can use semantic-preserving transformations, analogue to mutation operators, to improve the performance of defect detection tools in the testing stage. We first collect existing publications which implemented semantic-preserving transformations and share their implementation, such that we can reuse them. We empirically study the effectiveness of three different ensemble strategies for enhancing defect detection tools. We apply the collected transformations on the Devign dataset, considering vulnerabilities as a type of defect, and two fine-tuned LLMs for defect detection (VulBERTa, PLBART). We found 28 publications with 94 different transformations. We choose to implement 39 transformations from four of the publications, but a manual check revealed that 23 out 39 transformations change code semantics. Using the 16 remaining, correct transformations and three ensemble strategies, we were not able to increase the accuracy of the defect detection models. Our results show that reusing shared semantic-preserving transformation is difficult, sometimes even causing wrongful changes to the semantics. Keywords: defect detection, LLM, semantic-preserving transformation, ensemble

Collections

Summary

We haven't generated a summary for this paper yet.

Summarize Now

Paper Prompts

Explore 10 Community Prompts

Follow-up Questions

We haven't generated follow-up questions for this paper yet.

Generate Now

Semantic-Preserving Transformations as Mutation Operators: A Study on Their Effectiveness in Defect Detection (2503.23448v1)

Collections

Summary

Paper Prompts

Follow-up Questions

Authors (3)

Don't miss out on important new AI/ML research

Semantic-Preserving Transformations as Mutation Operators: A Study on Their Effectiveness in Defect Detection (2503.23448v1)

Collections

Summary

Paper Prompts

Follow-up Questions

Related Papers

Authors (3)

Don't miss out on important new AI/ML research