VLDB 2026 Research / reviewers in the wild / expert
Abir Bouraffa
dblp:276/2205
· DBLP profile ↗
8ranked-venue papers
2as first author
7since 2021 · last 2025
0000-0002-0980-625XORCID · corroborated
Domains — the database's venue-derived domains; a paper can count in several
Software engineering, systems software and programming languages · 8 · 2 first-author · 7 since 2021Databases, data management, data science and information retrieval · 1 · 1 since 2021
| Year | Publication | Venue | Position |
|---|---|---|---|
| 2025 | Not One to Rule Them All: Mining Meaningful Code Review Orders From GitHubabstractDevelopers use tools such as GitHub pull requests to review code, discuss proposed changes, and request modifications. While changed files are commonly presented in alphabetical order, this does not necessarily coincide with the reviewer’s preferred navigation sequence. This study investigates the different navigation orders developers follow while commenting on changes submitted in pull requests. We mined code review comments from 23,241 pull requests in 100 popular Java and Python repositories on GitHub to analyze the order in which the reviewers commented on the submitted changes. Our analysis shows that for 44.6% of pull requests, the reviewers comment in a non-alphabetical order. Among these pull requests, we identified traces of alternative meaningful orders: 20.6% (2,134) followed a largest-diff first order, 17.6% (1,827) were commented in the order of the files’ similarity to the pull request’s title and description, and 29% (1,188) of pull requests containing changes to both production and test files adhered to a test-first order. We also observed that the proportion of reviewed files to total submitted files was significantly higher in non-alphabetically ordered reviews, which also received slightly fewer approvals from reviewers, on average. Our findings highlight the need for additional support during code reviews, particularly for larger pull requests, where reviewers are more likely to adopt complex strategies rather than following a single predefined order. Abir Bouraffa, Carolin E. Brandt, Andy Zaidman, Walid Maalej |
EASE | 1 |
| 2025 | Explaining Explanations: An Empirical Study of Explanations in Code ReviewsabstractCode reviews are central for software quality assurance. Ideally, reviewers should explain their feedback to enable authors of code changes to understand the feedback and act accordingly. Different developers might need different explanations in different contexts. Therefore, assisting this process first requires understanding the types of explanations reviewers usually provide. The goal of this article is to study the types of explanations used in code reviews and explore the potential of Large Language Models (LLMs), specifically ChatGPT, in generating these specific types. We extracted 793 code review comments from Gerrit and manually labeled them based on whether they contained a suggestion, an explanation, or both. Our analysis shows that 42% of comments only include suggestions without explanations. We categorized the explanations into seven distinct types including rule or principle, similar examples, and future implications. When measuring their prevalence, we observed that some explanations are used differently by novice and experienced reviewers. Our manual evaluation shows that, when the explanation type is specified, ChatGPT can correctly generate the explanation in 88 out of 90 cases. This foundational work highlights the potential for future automation in code reviews, which can assist developers in sharing and obtaining different types of explanations as needed, thereby reducing back-and-forth communication. Ratnadira Widyasari, Ting Zhang 0011, Abir Bouraffa, Walid Maalej, David Lo 0001 |
ACM Trans. Softw. Eng. Methodol. | 3 |
| 2023 | Developers' Visuo-spatial Mental Model and Program ComprehensionabstractPrevious works from research and industry have proposed a spatial representation of code in a canvas, arguing that a navigational code space confers developers the freedom to organise elements according to their understanding. By allowing developers to translate logical relatedness into spatial proximity, this code representation could aid in code navigation and comprehension. However, the association between developers' code comprehension and their visuo-spatial mental model of the code is not yet well understood. This mental model is affected on the one hand by the spatial code representation and on the other by the visuo-spatial working memory of developers. We address this knowledge gap by conducting an online experiment with 20 developers following a between-subject design. The control group used a conventional tab-based code visualization, while the experimental group used a code canvas to complete three code comprehension tasks. Furthermore, we measure the participants' visuo-spatial working memory using a Corsi Block test at the end of the tasks. Our results suggest that, overall, neither the spatial representation of code nor the visuo-spatial working memory of developers has a significant impact on comprehension performance. However, we identified significant differences in the time dedicated to different comprehension activities such as navigation, annotation, and UI interactions. Abir Bouraffa, Gian-Luca Fuhrmann, Walid Maalej |
ICSE | 1 |
| 2022 | Beyond Duplicates: Towards Understanding and Predicting Link Types in Issue Tracking SystemsabstractSoftware projects use Issue Tracking Systems (ITS) like JIRA to track issues and organize the workflows around them. Issues are often inter-connected via different links such as the default JIRA link types Duplicate, Relate, Block, or Subtask. While previous research has mostly focused on analyzing and predicting duplication links, this work aims at understanding the various other link types, their prevalence, and characteristics towards a more reliable link type prediction. For this, we studied 607,208 links connecting 698,790 issues in 15 public JIRA repositories. Besides the default types, the custom types Depend, Incorporate, Split, and Cause were also common. We manually grouped all 75 link types used in the repositories into five general categories: General Relation, Duplication, Composition, Temporal / Causal, and Workflow. Comparing the structures of the corresponding graphs, we observed several trends. For instance, Duplication links tend to represent simpler issue graphs often with two components and Composition links present the highest amount of hierarchical tree structures (97.7%). Surprisingly, General Relation links have a significantly higher transitivity score than Duplication and Temporal / Causal links. Clara Marie Lüders, Abir Bouraffa, Walid Maalej |
MSR | 2 |
| 2022 | Empirical research on requirements quality: a systematic mapping studyabstractResearch has repeatedly shown that high-quality requirements are essential for the success of development projects. While the term "quality" is pervasive in the field of requirements engineering and while the body of research on requirements quality is large, there is no meta-study of the field that overviews and compares the concrete quality attributes addressed by the community. To fill this knowledge gap, we conducted a systematic mapping study of the scientific literature. We retrieved 6905 articles from six academic databases, which we filtered down to 105 relevant primary studies. The primary studies use empirical research to explicitly define, improve, or evaluate requirements quality. We found that empirical research on requirements quality focuses on improvement techniques, with very few primary studies addressing evidence-based definitions and evaluations of quality attributes. Among the 12 quality attributes identified, the most prominent in the field are ambiguity, completeness, consistency, and correctness. We identified 111 sub-types of quality attributes such as "template conformance" for consistency or "passive voice" for ambiguity. Ambiguity has the largest share of these sub-types. The artefacts being studied are mostly referred to in the broadest sense as "requirements", while little research targets quality attributes in specific types of requirements such as use cases or user stories. Our findings highlight the need to conduct more empirically grounded research defining requirements quality, using more varied research methods, and addressing a more diverse set of requirements types. Lloyd Montgomery, Davide Fucci, Abir Bouraffa, Lisa Scholz, Walid Maalej |
Requir. Eng. | 3 |
| 2022 | Correction to: Empirical research on requirements quality: a systematic mapping study
Lloyd Montgomery, Davide Fucci, Abir Bouraffa, Lisa Scholz, Walid Maalej |
Requir. Eng. | 3 |
| 2021 | The Role of Linguistic Relativity on the Identification of Sustainability Requirements: An Empirical StudyabstractLinguistic-Relativity-Theory states that language and its structure influence people’s world view and cognition. We investigate how this theory impacts the identification of requirements in practice. To this end, we conducted two controlled experiments with 101 participants. We randomly showed participants a set of requirements dimensions (i.e. a language structure) either with a focus on software quality or on sustainability and asked them to identify the requirements for a grocery shopping app according to these dimensions. Participants of the control group were not given any dimensions. The results show that the use of requirements dimensions significantly increases the number of identified requirements in comparison to the control group. Furthermore, participants who were given the sustainability dimensions identified more sustainability requirements. In follow up interviews with 16 practitioners, the interviewees reported benefits of the dimensions such as a holistic guidance but were also concerned about the customers acceptance. Furthermore, they stated challenges of implementing sustainability dimensions in the daily business but also suggested solutions like establishing sustainability as a common standard. Our study indicates that carefully structuring requirements engineering along sustainability dimensions can guide development teams towards considering and ensuring software sustainability. Yen Dieu Pham, Abir Bouraffa, Marleen Hillen, Walid Maalej |
RE | 2 |
| 2020 | ShapeRE: Towards a Multi-Dimensional Representation for Requirements of Sustainable SoftwareabstractIn this paper we introduce the preliminary design of a new framework named ShapeRE, which will address two problems potentially impeding the development of sustainable software. The first problem is the common differentiation between “functional” and “non-functional” requirements, which hides relevant sustainability requirements and hampers their identification. The second problem is the lack of developer-oriented representation approaches to ensure that all requirements relevant to sustainable software will be finally implemented. To address these two issues ShapeRE provides an alternative multi-dimensional approach for identifying requirements and a developer-oriented representation guideline. We suggest nine dimensions which are social, economic, environmental, technical, individual, purpose, design-aesthetics, integrative and legal. The design of our frame-work relies on insights from the fields of linguistics and empirical aesthetics, requirements engineering, and building architecture. We introduce the framework and discuss our evaluation strategy consisting of an online-experiment, a case study, and a survey. Yen Dieu Pham, Abir Bouraffa, Walid Maalej |
RE | 2 |