Search CORE

4 research outputs found

On the Evaluation of NLP-based Models for Software Engineering

Author: Ahmadabadi Matin Nili
Izadi Maliheh
Publication venue: 'Association for Computing Machinery (ACM)'
Publication date: 01/01/2022
Field of study

NLP-based models have been increasingly incorporated to address SE problems. These models are either employed in the SE domain with little to no change, or they are greatly tailored to source code and its unique characteristics. Many of these approaches are considered to be outperforming or complementing existing solutions. However, an important question arises here: "Are these models evaluated fairly and consistently in the SE community?". To answer this question, we reviewed how NLP-based models for SE problems are being evaluated by researchers. The findings indicate that currently there is no consistent and widely-accepted protocol for the evaluation of these models. While different aspects of the same task are being assessed in different studies, metrics are defined based on custom choices, rather than a system, and finally, answers are collected and interpreted case by case. Consequently, there is a dire need to provide a methodological way of evaluating NLP-based models to have a consistent assessment and preserve the possibility of fair and efficient comparison.Comment: To appear in the Proceedings of the 1sth International Workshop on Natural Language-based Software Engineering (NLBSE), co-located with ICSE, 202

arXiv.org e-Print Archive

TU Delft Repository

A Fine-grained Data Set and Analysis of Tangling in Bug Fixing Commits

Author: Aghamohammadi Alireza
Ahmadabadi Matin Nili
Aktas Ethem Utku
Alam Omar
Albrecht Ella
Aldaeej Abdullah
Amit Idan
Bossenmaier Tim
Chahal Kuljit Kaur
Chakroborti Debasish
Colomo-Palacios Ricardo
Davis James
Davis Willard
Eismann Simon
Erbel Johannes
Fard Fatemeh
Ghaleb Taher Ahmed
Henley Austin Z.
Herbold Steffen
Hoy Nathaniel
Kourtzanidis Stratos
Ledel Benjamin
Lenarduzzi Valentina
Madeja Matej
Makedonski Philip
Malavolta Ivano
Marcilio Diego
Nagaria Bhaveet
Pashchenko Ivan
Qin Yihao
Rodríguez-Pérez Gema
Serebrenik Alexander
Shamasbi Simin Maleki
Singh Paramvir
Spieker Helge
Strüber Daniel
Sulir Matus
Szabados Kristof
Trautsch Alexander
Treude Christoph
Turhan Burak
Tuzun Eray
Verdecchia Roberto
Walunj Vijay
Wang Shangwen
Wickert Anna-Katharina
Wu Hongjun
Wyrich Marvin
Publication venue
Publication date: 01/01/2021
Field of study

Context: Tangled commits are changes to software that address multiple concerns at once. For researchers interested in bugs, tangled commits mean that they actually study not only bugs, but also other concerns irrelevant for the study of bugs. Objective: We want to improve our understanding of the prevalence of tangling and the types of changes that are tangled within bug fixing commits. Methods: We use a crowd sourcing approach for manual labeling to validate which changes contribute to bug fixes for each line in bug fixing commits. Each line is labeled by four participants. If at least three participants agree on the same label, we have consensus. Results: We estimate that between 17% and 32% of all changes in bug fixing commits modify the source code to fix the underlying problem. However, when we only consider changes to the production code files this ratio increases to 66% to 87%. We find that about 11% of lines are hard to label leading to active disagreements between participants. Due to confirmed tangling and the uncertainty in our data, we estimate that 3% to 47% of data is noisy without manual untangling, depending on the use case. Conclusion: Tangled commits have a high prevalence in bug fixes and can lead to a large amount of noise in the data. Prior research indicates that this noise may alter results. As researchers, we should be skeptics and assume that unvalidated data is likely very noisy, until proven otherwise.Comment: Status: Accepted at Empirical Software Engineerin

arXiv.org e-Print Archive

University of Oulu Repository - Jultika

Monash University Research Portal

A fine-grained data set and analysis of tangling in bug fixing commits

Author: AGHAMOHAMMADI Alireza
AHMADABADI Matin Nili
BOSSENMAIER Tim
COLOMO-PALACIOS Ricardo
GHALEB Taher Ahmed
HERBOLD Steffen
HOY Nathaniel G.
KAUR CHAHAL Kuljit
LEDEL Benjamin
MADEJA Matej
MAKEDONSKI Philip
NAGARIA Bhaveet
RODRÍGUEZ-PÉREZ Gema
SINGH Paramvir
SPIEKER Helge
SZABADOS Kristóf
TRAUTSCH Alexander
TREUDE Christoph
VERDECCHIA Roberto
WANG Shangwen
Publication venue: Springer
Publication date: 01/11/2022
Field of study

Institutional Knowledge at Singapore Management University

A fine-grained data set and analysis of tangling in bug fixing commits

Author: Aghamohammadi A. (Alireza)
Ahmadabadi M. N. (Matin Nili)
Aktas E. U. (Ethem Utku)
Alam O. (Omar)
Albrecht E. (Ella)
Aldaeej A. (Abdullah)
Amit I. (Idan)
Bossenmaier T. (Tim)
Chahal K. K. (Kuljit Kaur)
Chakroborti D. (Debasish)
Colomo-Palacios R. (Ricardo)
Davis J. (James)
Davis W. (Willard)
Eismann S. (Simon)
Erbel J. (Johannes)
Fard F. (Fatemeh)
Ghaleb T. A. (Taher A.)
Henley A. Z. (Austin Z.)
Herbold S. (Steffen)
Hoy N. (Nathaniel)
Kourtzanidis S. (Stratos)
Ledel B. (Benjamin)
Lenarduzzi V. (Valentina)
Madeja M. (Matej)
Makedonski P. (Philip)
Malavolta I. (Ivano)
Marcilio D. (Diego)
Nagaria B. (Bhaveet)
Pashchenko I. (Ivan)
Qin Y. (Yihao)
Rodríguez-Pérez G. (Gema)
Serebrenik A. (Alexander)
Shamasbi S. M. (Simin Maleki)
Singh P. (Paramvir)
Spieker H. (Helge)
Strüber D. (Daniel)
Sulír M. (Matúš)
Szabados K. (Kristof)
Trautsch A. (Alexander)
Treude C. (Christoph)
Turhan B. (Burak)
Tuzun E. (Eray)
Verdecchia R. (Roberto)
Walunj V. (Vijay)
Wang S. (Shangwen)
Wickert A.-K. (Anna-Katharina)
Wu H. (Hongjun)
Wyrich M. (Marvin)
Publication venue: Springer Nature
Publication date: 01/01/2022
Field of study

Abstract Context: Tangled commits are changes to software that address multiple concerns at once. For researchers interested in bugs, tangled commits mean that they actually study not only bugs, but also other concerns irrelevant for the study of bugs. Objectives: We want to improve our understanding of the prevalence of tangling and the types of changes that are tangled within bug fixing commits. Methods: We use a crowd sourcing approach for manual labeling to validate which changes contribute to bug fixes for each line in bug fixing commits. Each line is labeled by four participants. If at least three participants agree on the same label, we have consensus. Results: We estimate that between 17% and 32% of all changes in bug fixing commits modify the source code to fix the underlying problem. However, when we only consider changes to the production code files this ratio increases to 66% to 87%. We find that about 11% of lines are hard to label leading to active disagreements between participants. Due to confirmed tangling and the uncertainty in our data, we estimate that 3% to 47% of data is noisy without manual untangling, depending on the use case. Conclusions: Tangled commits have a high prevalence in bug fixes and can lead to a large amount of noise in the data. Prior research indicates that this noise may alter results. As researchers, we should be skeptics and assume that unvalidated data is likely very noisy, until proven otherwise

University of Oulu Repository - Jultika