Predicting unstable software benchmarks using static source code features

Laaber, Christoph; Basmaci, Mikael; Salza, Pasquale

doi:10.1007/s10664-021-09996-y

Cited by 18 publications

(2 citation statements)

References 118 publications

Supporting

Mentioning

Contrasting

Order By: Relevance

“…We chose 2014 as a cutoff point because this was the year the TensorFlow system was initially released. I2: Making use of DL as a core contribution of the paper and explicitly reporting on the used code representation approach. To illustrate this criterion, we discuss the following study as a counterexample [26]. In this study, Laaber et al.…”

Section: Methodsmentioning

confidence: 99%

“…� I2: Making use of DL as a core contribution of the paper and explicitly reporting on the used code representation approach. To illustrate this criterion, we discuss the following study as a counterexample [26]. In this study, Laaber et al tackled an SE task (predictability of system performance) and the authors used an artificial neural network (ANN) as a DL model for that task.…”

Section: Inclusion Criteriamentioning

confidence: 99%

See 1 more Smart Citation

A systematic mapping study of source code representation for deep learning in software engineering

et al. 2022

Self Cite

View full text Add to dashboard Cite

The usage of deep learning (DL) approaches for software engineering has attracted much attention, particularly in source code modelling and analysis. However, in order to use DL, source code needs to be formatted to fit the expected input form of DL models. This problem is known as source code representation. Source code can be represented via different approaches, most importantly, the tree-based, token-based, and graph-based approaches. We use a systematic mapping study to investigate i detail the representation approaches adopted in 103 studies that use DL in the context of software engineering. Thus, studies are collected from 2014 to 2021 from 14 different journals and 27 conferences. We show that each way of representing source code can provide a different, yet orthogonal view of the same source code. Thus, different software engineering tasks might require different (combinations of) code representation approaches, depending on the nature and complexity of the task. Particularly, we show that it is crucial to define whether the DL approach requires lexical, syntactical, or semantic code information. Our analysis shows that a wide range of different representations and combinations of representations (hybrid representations) are used to solve a wide range of common software engineering problems. However, we also observe that current research does not generally attempt to transfer existing representations or models to other studies even though there are other contexts in which these representations and models may also be useful. We believe that there is potential for more reuse and the application of transfer learning when applying DL to software engineering tasks.

show abstract

Section: Methodsmentioning

confidence: 99%

Section: Inclusion Criteriamentioning

confidence: 99%