Krzysztof Stencel

dblp:s/KrzysztofStencel · DBLP profile ↗
← Back
19ranked-venue papers in the field
0as first author
7since 2021 · last 2025
0000-0001-6356-4872ORCID · verified

Domains — venue-derived; a paper can count in several

Database Systems & Data Management · 9Other / Interdisciplinary · 6Big Data, Cloud & Distributed Data Systems · 2Knowledge Engineering, Semantic Web & Information Systems · 2
YearPublicationVenuePosition
2025 Out of Sight, Still at Risk: The Lifecycle of Transitive Vulnerabilities in Maven
abstract
The modern software development landscape heavily relies on transitive dependencies. They enable seamless integration of third-party libraries. However, they also introduce security challenges. Transitive vulnerabilities that arise from indirect dependencies expose projects to risks associated with Common Vulnerabilities and Exposures (CVEs). It happens even when direct dependencies remain secure. This paper examines the lifecycle of transitive vulnerabilities in the Maven ecosystem. We employ survival analysis to measure the time projects remain exposed after a CVE is introduced. Using a large dataset of Maven projects, we identify factors that influence the resolution of these vulnerabilities. Our findings offer practical advice on improving dependency management.
Piotr Przymus, Mikolaj Fejzer, Jakub Narebski, Krzysztof Rykaczewski, Krzysztof Stencel
MSR5
2025 HaPy-Bug - Human Annotated Python Bug Resolution Dataset
abstract
We present HaPy-Bug, a curated dataset of 793 Python source code commits associated with bug fixes, with each line of code annotated by three domain experts. The annotations offer insights into the purpose of modified files, changes at the line level, and reviewers’ confidence levels. We analyze HaPy-Bug to examine the distribution of file purposes, types of modifications, and tangled changes. Additionally, we explore its potential applications in bug tracking, the analysis of bug-fixing practices, and the development of repository analysis tools. HaPy-Bug serves as a valuable resource for advancing research in software maintenance and security.
Piotr Przymus, Mikolaj Fejzer, Jakub Narebski, Radoslaw Wozniak, Lukasz Halada, Aleksander Kazecki, Mykhailo Molchanov, Krzysztof Stencel
MSR8
2024 Recursive Queries: Twenty-Five Years After SQL: 1999
abstract
SQL:1999 recursive queries are almost a quarter century old. In this standard the recursive queries have the form of recursive common table expressions. In recent years vendors of almost all database systems finally implemented this feature. In this article we analyze those implementations and propose a benchmark that can be used to measure them in the future. We also present an experimental evaluation of recursive SQL:1999 queries in free-to-use popular database systems using a network (graph) data set and a hierarchical (tree) data set.
Marta Burzanska, Piotr Wisniewski 0001, Krzysztof Stencel
IEEE Big Data3
2024 How I Learned to Stop Worrying and Love ChatGPT
abstract
In the dynamic landscape of software engineering, the emergence of ChatGPT-generated code signifies a distinctive and evolving paradigm in development practices. We delve into the impact of interactions with ChatGPT on the software development process, specifically analysing its influence on source code changes. Our emphasis lies in aligning code with ChatGPT conversations, separately analysing the user-provided context of the code and the extent to which the resulting code has been influenced by ChatGPT. Additionally, employing survival analysis techniques, we examine the longevity of ChatGPT-generated code segments in comparison to lines written traditionally. The goal is to provide valuable insights into the transformative role of ChatGPT in software development, illuminating its implications for code evolution and sustainability within the ecosystem.
Piotr Przymus, Mikolaj Fejzer, Jakub Narebski, Krzysztof Stencel
MSR4
2023 The Secret Life of CVEs
abstract
The Common Vulnerabilities and Exposures (CVEs) system is a reference method for documenting publicly known information security weaknesses and exposures. This paper presents a study of the lifetime of CVEs in software projects and the risk factors affecting their existence. The study uses survival analysis to examine how features of programming languages, projects, and CVEs themselves impact the lifetime of CVEs. We suggest avenues for future research to investigate the effect of various factors on the resolution of vulnerabilities.
Piotr Przymus, Mikolaj Fejzer, Jakub Narebski, Krzysztof Stencel
MSR4
2023 A practical study of methods for deriving insightful attribute importance rankings using decision bireducts
abstract
Subject matter experts (SMEs) often rely on attribute importance rankings to verify machine learning models, acquire insights into their outcomes, and gain a deeper understanding of the investigated phenomena. To further increase their usefulness, we introduce a new approach to the evaluation of attribute rankings produced by any machine learning method . As a real-world case study , we investigate the attribute importance scores produced using XGBoost and decision bireducts on the data gathered by an HR company, where the goal is to predict the willingness of candidates to change their job. For this task, XGBoost delivers accurate models but fails to identify many attributes that are important to SMEs. In comparison, decision bireducts lead to models that are easier to interpret and explore the data with a higher focus on the diversity of attributes. The ensembles of decision bireducts deliver comparable accuracy and their associated attribute rankings are more insightful than those of XGBoost.
Andrzej Janusz, Dominik Slezak, Sebastian Stawicki, Krzysztof Stencel
Inf. Sci.4
2022 Large Scale Windowed Matching
abstract
Missing or invalid records in sales data are a common obstacle that can damage the overall effectiveness of market analysis. Completing the data on the basis of the records obtained so far can be formulated in means of a schema matching task. In this paper we present a machine learning based method for performing schema matching for transactional data. The analysis is based on a dataset of over 700.000 transactions from retail stores. We confront the proposed solution with manual and conventional approaches.
Mateusz Przyborowski, Krzysztof Ciebiera, Krzysztof Stencel
IEEE Big Data3
2018 How to Match Jobs and Candidates - A Recruitment Support System Based on Feature Engineering and Advanced Analytics
Andrzej Janusz, Sebastian Stawicki, Michal Drewniak, Krzysztof Ciebiera, Dominik Slezak, Krzysztof Stencel
IPMU (2)6
2018 SENSEI: An Intelligent Advisory System for the eSport Community and Casual Players
abstract
In this article, we describe the SENSEI system. It helps players to improve their skills in popular eSports games. We discuss the main goals of the system and explain the associated challenges. We also present its conceptual architecture which aims at enabling full automation of the data acquisition and analytic processes. The system is expected to provide in-depth analytics of players' performance and give practical advice regarding possible improvements. Thus its architecture allows players to provide feedback and manually label important concepts. Finally, we discuss our first case study - an advisory system for popular collectible card video games.
Andrzej Janusz, Dominik Slezak, Sebastian Stawicki, Krzysztof Stencel
WI4
2018 Profile based recommendation of code reviewers
abstract
Code reviews consist in proof-reading proposed code changes in order to find their shortcomings such as bugs, insufficient test coverage or misused design patterns. Code reviews are conducted before merging submitted changes into the main development branch. The selection of suitable reviewers is crucial to obtain the high quality of reviews. In this article we present a new method of recommending reviewers for code changes. This method is based on profiles of individual programmers. For each developer we maintain his/her profile. It is the multiset of all file path segments from commits reviewed by him/her. It will get updated when he/she presents a new review. We employ a similarity function between such profiles and change proposals to be reviewed. The programmer whose profile matches the change most is recommended to become the reviewer. We performed an experimental comparison of our method against state-of-the-art techniques using four large open-source projects. We obtained improved results in terms of classification metrics (precision, recall and F-measure) and performance (we have lower time and space complexity).
Mikolaj Fejzer, Piotr Przymus, Krzysztof Stencel
J. Intell. Inf. Syst.3
2015 Translating Relational Queries into Spreadsheets
abstract
Spreadsheets are among the most commonly used applications for data management and analysis. They combine data processing with very diverse supplementary features: statistics, visualization, reporting, linear programming solvers, Web queries periodically downloading data from external sources, etc. However, the spreadsheet paradigm of computation still lacks sufficient analysis. In this article, we demonstrate that a spreadsheet can implement all data transformations definable in SQL, merely by utilizing spreadsheet formulas. We provide a query compiler, which translates any given SQL query into a worksheet of the same semantics, including NULL values. Thereby, database operations become available to the users who do not want to migrate to a database. They can define their queries using a high-level language and then get their execution plans in a plain vanilla spreadsheet. The functions available in spreadsheets impose limitations on the algorithms one can implement. In this paper, we offer O(n log2n) sorting spreadsheet, using a non-constant number of rows, and, surprisingly, Depth-First-Search and Breadth-First-Search on graphs.
Jacek Sroka, Adrian Panasiuk, Krzysztof Stencel, Jerzy Tyszkiewicz
IEEE Trans. Knowl. Data Eng.3
2014 Open Source Is a Continual Bugfixing by a Few
Mikolaj Fejzer, Michal Wojtyna, Marta Burzanska, Piotr Wisniewski 0001, Krzysztof Stencel
ADBIS5
2013 On Materializing Paths for Faster Recursive Querying
Aleksandra Boniewicz, Piotr Wisniewski 0001, Krzysztof Stencel
ADBIS (2)3
2012 Optimization of Object-Oriented Queries through Pushing Selections
Marcin Drozd, Michal Bleja, Krzysztof Stencel, Kazimierz Subieta
ADBIS (2)3
2012 PropScale: An Update Propagator for Joint Scalable Storage
Pawel Leszczynski, Krzysztof Stencel
ADBIS (2)2
2012 Extending HQL with Plain Recursive Facilities
Aneta Szumowska, Marta Burzanska, Piotr Wisniewski 0001, Krzysztof Stencel
ADBIS (2)4
2010 Consistent Caching of Data Objects in Database Driven Websites
Pawel Leszczynski, Krzysztof Stencel
ADBIS2
2009 Pushing Predicates into Recursive SQL Common Table Expressions
Marta Burzanska, Krzysztof Stencel, Piotr Wisniewski 0001
ADBIS2
2005 Usable Recursive Queries
Tomasz Pieciukiewicz, Krzysztof Stencel, Kazimierz Subieta
ADBIS2