Thomas Cerqueus

dblp:11/7834 · DBLP profile ↗
← Back
6ranked-venue papers
2as first author
0since 2021 · last 2017
0000-0003-3484-188XORCID · corroborated

Domains — the database's venue-derived domains; a paper can count in several

Databases, data management, data science and information retrieval · 5 · 1 first-authorArtificial intelligence and machine learning · 3 · 1 first-authorComputer networks · 1 · 1 first-author

Expertise — from the expertise taxonomy: the topics of the expert's papers under the CCF categories. A weight counts papers with recency: 1 for a paper about the topic, 0.3 when the topic is its context, halved every five years.

Software engineering, system software, and programming languages
1 paper
Software maintenance and evolution · 100%

Topics — the 2 heaviest of 3, each with the papers that count most for it

TopicWeightPapersLastEvidence papers
Software maintenance and evolution › software evolution
schema evolution
0.212015
ControVol: A framework for controlled schema evolution in NoSQL application development · ICDE 2015
Software maintenance and evolution
integrated development environment support
0.112015
ControVol: A framework for controlled schema evolution in NoSQL application development · ICDE 2015

Methods — techniques the papers use, named apart from their topics

static type checking · 0.2object mapper · 0.2
YearPublicationVenuePosition
2017 ReX: Representative extrapolating relational databases
Teodora Sandra Buda, Thomas Cerqueus, Cristian Grava, John Murphy 0001
Inf. Syst.2
2015 ControVol: A framework for controlled schema evolution in NoSQL application development
abstract
Building scalable web applications on top of NoSQL data stores is becoming common practice. Many of these data stores can easily be accessed programmatically, and do not enforce a schema. Software engineers can design the data model on the go, a flexibility that is crucial in agile software development. The typical tasks of database schema management are now handled within the application code, usually involving object mapper libraries. However, today's Integrated Development Environments (IDEs) lack the proper tool support when it comes to managing the combined evolution of the application code and of the schema. Yet simple refactorings such as renaming an attribute at the source code level can cause irretrievable data loss or runtime errors once the application is serving in production. In this demo, we present ControVol, a framework for controlled schema evolution in application development against NoSQL data stores. ControVol is integrated into the IDE and statically type checks object mapper class declarations against the schema evolution history, as recorded by the code repository. ControVol is capable of warning of common yet risky cases of mismatched data and schema. ControVol is further able to suggest quick fixes by which developers can have these issues automatically resolved.
Stefanie Scherzinger, Thomas Cerqueus, Eduardo C. de Almeida
ICDE2
2014 VFDS: An Application to Generate Fast Sample Databases
abstract
Large amounts of data often require expensive and time-consuming analysis. Therefore, highly scalable and efficient techniques are necessary to process, analyze and discover useful information. Database sampling has proven to be a powerful method to surpass these limitations. Using only a sample of the original large database brings the benefit of obtaining useful information faster, at the potential expense of lower accuracy. In this paper, we demonstrate \vfds, a novel fast database sampling system that maintains the referential integrity of the data. The system is developed over the open-source database management system, MySQL. We present various scenarios to demonstrate the effectiveness of VFDS in approximate query answering, sample size, and execution time, on both real and synthetic databases.
Teodora Sandra Buda, Thomas Cerqueus, John Murphy 0001, Morten Kristiansen
CIKM2
2013 CoDS: A Representative Sampling Method for Relational Databases
Teodora Sandra Buda, Thomas Cerqueus, John Murphy 0001, Morten Kristiansen
DEXA (1)2
2012 An approach to manage semantic heterogeneity in unstructured P2P information retrieval systems
abstract
In unstructured information retrieval P2P systems, semantic heterogeneity comes from the use of different ontologies. Semantic interoperability refers to the ability of peers to communicate with each others. We take into account these notions separately, as raising two different problems. Hence we propose two independent and complementary solutions. The GoOD-TA protocol aims at reducing heterogeneity through ontology-driven topology adaptation. DiQuESh is a top-k algorithm for distributed information retrieval that is intended to ensure interoperability. This distinction enables highlighting their respective benefits on the IR performances and leads to a modular architecture. For our experiments we obtained a set of actively used real-world ontologies through the NCBO BioPortal. We implemented GoOD-TA and DiQuESH in Java and used the PeerSim simulator. We first show that GoOD-TA nicely reduces the semantic heterogeneity related to the system topology, handles the evolution of peers' descriptors, and is suitable for dynamic systems. Then, GoOD-TA and DiQuESh are run simultaneously, with a significant increase of precision and recall. This enables to identify the indirect contribution of heterogeneity reduction obtained with GoOD-TA to improving interoperability.
Thomas Cerqueus, Sylvie Cazalens, Philippe Lamarre
P2P1
2011 Semantic Heterogeneity Measures of Unstructured P2P Systems
abstract
We consider P2P data sharing systems in which each participant uses an ontology to represent information. If all the partipants do not use the same ontology, the system is said to be semantically heterogeneous. Several methods have been proposed to reach a degree of interoperability but thorough evaluation of these methods is prevented by a lack of tools to describe the situations in which they have been tested. In this paper we identify components that impact on the semantic heterogeneneity, and we define several complementary measures to capture the different facets of heterogeneity. Proposed measures allow to characterize the situation in which a method is evaluated, or to measure the heterogeneity reduction produced by another method.
Thomas Cerqueus, Sylvie Cazalens, Philippe Lamarre
Web Intelligence1