EDBT 2026 Demo / reviewers in the wild / expert
Diego Calvanese
dblp:c/DiegoCalvanese
· DBLP profile ↗
66ranked-venue papers in the field
31as first author
13since 2021 · last 2026
0000-0001-5174-9693ORCID · verified
Domains — venue-derived; a paper can count in several
Database Systems & Data Management · 32 (20 first)Knowledge Engineering, Semantic Web & Information Systems · 18 (5 first)Business Process & Enterprise Data · 11 (4 first)Information Retrieval & Web Search · 3 (2 first)Big Data, Cloud & Distributed Data Systems · 1Other / Interdisciplinary · 1
| Year | Publication | Venue | Position |
|---|---|---|---|
| 2026 | News-Informed Probabilistic Models for AI Risk Analysis
Mattia Fumagalli, Stefano M. Nicoletti, Diego Calvanese, Giancarlo Guizzardi |
CAiSE (2) | 3 |
| 2026 | CMiner: An Algorithm to Discover Frequent Structures in Conceptual Models
Simone Avellino, Emanuele Valore, Giovanni Micale, Antonio Di Maria, Mattia Fumagalli, Tiago Prince Sales, Alfredo Pulvirenti, Diego Calvanese |
EDBT | 8 |
| 2026 | Agentic Business Process Management: A research manifestoabstractThis paper presents a manifesto that articulates the conceptual foundations of Agentic Business Process Management (APM), an extension of Business Process Management (BPM) for governing autonomous agents executing processes in organizations. From a management perspective, APM represents a paradigm shift from the traditional view on business processes. This shift is driven by the realization of process awareness by agent-oriented abstractions: software and human agents act as primary functional entities that perceive, reason, and act within explicit process frames. Thus, APM moves away from automation-oriented BPM towards systems in which autonomy is constrained, aligned, and made operational through process aware agents. We introduce the core abstractions and architectural elements required to realize APM systems and elaborate on four key capabilities that agents in APM systems must support: framed autonomy , explainability , conversational actionability , and self-modification . These capabilities jointly ensure that agents’ goals are aligned with organizational goals and that agents behave in a framed yet proactive manner in pursuing those goals. We discuss the extent to which the capabilities can be realized and identify research challenges whose resolution requires further advances in BPM, AI, and multi-agent systems. The manifesto thus serves as a roadmap for bridging these communities and for guiding the development of APM systems in practice. Diego Calvanese, Angelo Casciani, Giuseppe De Giacomo, Marlon Dumas, Fabiana Fournier, Timotheus Kampik, Emanuele La Malfa, Lior Limonad, Andrea Marrella, Andreas Metzger, Marco Montali, Daniel Amyot, Peter Fettke, Artem Polyvyanyy, Stefanie Rinderle-Ma, Sebastian Sardiña, Niek Tax, Barbara Weber |
Inf. Syst. | 1 |
| 2025 | Compact Answers to Temporal Path Queries
Diego Calvanese, Julien Corman, Anton Dignös, Werner Nutt, Ognjen Savkovic |
ISWC (1) | 2 |
| 2025 | Virtual Knowledge Graphs over Earth Observation Data
Albulen Pano, Davide Lanti, Diego Calvanese |
ISWC (2) | 3 |
| 2024 | Evaluating quality of ontology-driven conceptual models abstractionsabstractThe complexity of an (ontology-driven) conceptual model highly correlates with the complexity of the domain and software for which it is designed. With that in mind, an algorithm for producing ontology-driven conceptual model abstractions was previously proposed. In this paper, we empirically evaluate the quality of the abstractions produced by it. First, we have implemented and tested the last version of the algorithm over a FAIR catalog of models represented in the ontology-driven conceptual modeling language OntoUML. Second, we performed three user studies to evaluate the usefulness of the resulting abstractions as perceived by modelers. This paper reports on the findings of these experiments and reflects on how they can be exploited to improve the existing algorithm. Elena Romanenko, Diego Calvanese, Giancarlo Guizzardi |
Data Knowl. Eng. | 2 |
| 2024 | Verification of Unary Communicating Datalog ProgramsabstractWe study verification of reachability properties over Communicating Datalog Programs (CDPs), which are networks of relational nodes connected through unordered channels and running Datalog-like computations. Each node manipulates a local state database (DB), depending on incoming messages and additional input DBs from external services. Decidability of verification for CDPs has so far been established only under boundedness assumptions on the state and channel sizes, showing at the same time undecidability of reachability for unbounded states with only two unary relations or unbounded channels with a single binary relation. The goal of this paper is to study the open case of CDPs with bounded states and unbounded channels, under the assumption that channels carry unary relations only. We discuss the significance of the resulting model and prove the decidability of verification of variants of reachability, captured in fragments of first-order CTL. We do so through a novel reduction to coverability problems in a class of high-level Petri Nets that manipulate unordered data identifiers. We study the tightness of our results, showing that minor generalizations of the considered reachability properties yield undecidability of verification, both for CDPs and the corresponding Petri Net model. C. Aiswarya, Diego Calvanese, Francesco Di Cosmo, Marco Montali |
Proc. ACM Manag. Data | 2 |
| 2023 | Extracting Event Data from Document-Driven Enterprise Systems
Diego Calvanese, Mieke Jans, Tahir Emre Kalayci, Marco Montali |
CAiSE | 1 |
| 2023 | Conceptually-grounded mapping patterns for Virtual Knowledge GraphsabstractVirtual Knowledge Graphs (VKGs) constitute one of the most promising paradigms for integrating and accessing legacy data sources. A critical bottleneck in the integration process involves the definition, validation, and maintenance of mapping assertions that link data sources to a domain ontology. To support the management of mappings throughout their entire lifecycle, we identify a comprehensive catalog of sophisticated mapping patterns that emerge when linking databases to ontologies. To do so, we build on well-established methodologies and patterns studied in data management, data analysis, and conceptual modeling. These are extended and refined through the analysis of concrete VKG benchmarks and real-world use cases, and considering the inherent impedance mismatch between data sources and ontologies. We validate our catalog on the considered VKG scenarios, showing that it covers the vast majority of mappings present therein. Diego Calvanese, Avigdor Gal, Davide Lanti, Marco Montali, Alessandro Mosca 0001, Roee Shraga |
Data Knowl. Eng. | 1 |
| 2022 | Towards Pragmatic Explanations for Domain Ontologies
Elena Romanenko, Diego Calvanese, Giancarlo Guizzardi |
EKAW | 2 |
| 2021 | ADaMaP: Automatic Alignment of Relational Data Sources Using Mapping Patterns
Diego Calvanese, Avigdor Gal, Naor Haba, Davide Lanti, Marco Montali, Alessandro Mosca 0001, Roee Shraga |
CAiSE | 1 |
| 2021 | Consistency assessment for open geodata integration: an ontology-based approach
Linfang Ding, Guohui Xiao 0001, Diego Calvanese, Liqiu Meng |
GeoInformatica | 3 |
| 2021 | Towards the next generation of the LinkedGeoData project using virtual knowledge graphsabstractWith the advancement of Semantic Technologies, large geospatial data sources have been increasingly published as Linked data on the Web. The LinkedGeoData project is one of the most prominent such projects to create a large knowledge graph from OpenStreetMap (OSM) with global coverage and interlinking of other data sources. In this paper, we report on the ongoing effort of exposing the relational database in LinkedGeoData as a SPARQL endpoint using Virtual Knowledge Graph (VKG) technology. Specifically, we present two realizations of VKGs, using the two systems Sparqlify and Ontop. In order to improve compliance with the OGC GeoSPARQL standard, we have implemented GeoSPARQL support in Ontop v4. Moreover, we have evaluated the VKG-powered LinkedGeoData in the test areas of Italy and Germany. Our experiments demonstrate that such system supports complex GeoSPARQL queries, which confirms that query answering in the VKG approach is efficient. Linfang Ding, Guohui Xiao 0001, Albulen Pano, Claus Stadler, Diego Calvanese |
J. Web Semant. | 5 |
| 2020 | Semantic Integration of Bosch Manufacturing Data Using Virtual Knowledge Graphs
Elem Guzel Kalayci, Irlán Grangel-González, Felix Lösch, Guohui Xiao 0001, Anees Mehdi, Evgeny Kharlamov, Diego Calvanese |
ISWC (2) | 7 |
| 2020 | The Virtual Knowledge Graph System Ontop
Guohui Xiao 0001, Davide Lanti, Roman Kontchakov, Sarah Komla-Ebri, Elem Guzel Kalayci, Linfang Ding, Julien Corman, Benjamin Cogrel, Diego Calvanese, Elena Botoeva |
ISWC (2) | 9 |
| 2019 | Modeling and In-Database Management of Relational, Data-Aware Processes
Diego Calvanese, Marco Montali, Fabio Patrizi, Andrey Rivkin |
CAiSE | 1 |
| 2019 | On expansion and contraction of DL-Lite knowledge bases
Dmitriy Zheleznyakov, Evgeny Kharlamov, Werner Nutt, Diego Calvanese |
J. Web Semant. | 4 |
| 2018 | Semantic Technologies for Data Access and IntegrationabstractRecently, semantic technologies have been successfully deployed to overcome the typical difficulties in accessing and integrating data stored in different kinds of legacy sources. In particular, knowledge graphs are being used as a mechanism to provide a uniform representation of heterogeneous information. Such graphs represent data in the RDF format, which is complemented by an ontology and can be queried using the standard SPARQL language. The RDF graph is often obtained by materializing source data, following the traditional extract-transform-load workflow. Alternatively, the sources are declaratively mapped to the ontology, and the RDF graph is maintained virtual. In such an approach, usually called ontology-based data access/integration (OBDA/I), query answering is based on sophisticated query transformation techniques. In this tutorial: (i) we provide a general introduction to semantic technologies; (ii) we illustrate the principles underlying OBDA/I, providing insights into its theoretical foundations, and describing well-established algorithms, techniques, and tools; (iii) we discuss relevant use-cases for OBDA/I; (iv) we provide an overview on some recent advancements. Diego Calvanese, Guohui Xiao 0001 |
CIKM | 1 |
| 2018 | Ontop-temporal: A Tool for Ontology-based Query Answering over Temporal DataabstractWe present Ontop-temporal, an extension of the ontology-based data access system Ontop for query answering with temporal data and ontologies. Ontop is a system to answer SPARQL queries over various data stores, using standard R2RML mappings and an OWL2QL domain ontology to produce high-level conceptual views over the raw data. The Ontop-temporal extension is designed to handle timestamped log data, by additionally using (i) mappings supporting validity time specification, and (ii) rules based on metric temporal logic to define temporalised concepts. In this demo we present how Ontop-temporal can be used to facilitate the access to the MIMIC-III critical care unit dataset containing log data on hospital admissions, procedures, and diagnoses. We use the ICD9CM diagnoses ontology and temporal rules formalising the selection of patients for clinical trials taken from the clinicaltrials.gov database. We demonstrate how high-level queries can be answered by Ontop-temporal to identify patients eligible for the trials. Elem Guzel Kalayci, Guohui Xiao 0001, Vladislav Ryzhikov, Tahir Emre Kalayci, Diego Calvanese |
CIKM | 5 |
| 2018 | Conceptual Schema Transformation in Ontology-Based Data Access
Diego Calvanese, Tahir Emre Kalayci, Marco Montali, Ario Santoso, Wil M. P. van der Aalst |
EKAW | 1 |
| 2018 | Efficient Ontology-Based Data Integration with Canonical IRIs
Guohui Xiao 0001, Dag Hovland, Dimitris Bilidas, Martín Rezk, Martin Giese, Diego Calvanese |
ESWC | 6 |
| 2018 | Expressivity and Complexity of MongoDB QueriesabstractA significant number of novel database architectures and data models have been proposed during the last decade. While some of these new systems have gained in popularity, they lack a proper formalization, and a precise understanding of the expressivity and the computational properties of the associated query languages. In this paper, we aim at filling this gap, and we do so by considering MongoDB, a widely adopted document database managing complex (tree structured) values represented in a JSON-based data model, equipped with a powerful query mechanism. We provide a formalization of the MongoDB data model, and of a core fragment, called MQuery, of the MongoDB query language. We study the expressivity of MQuery, showing its equivalence with nested relational algebra. We further investigate the computational complexity of significant fragments of it, obtaining several (tight) bounds in combined complexity, which range from LOGSPACE to alternating exponential-time with a polynomial number of alternations. As a consequence, we obtain also a characterization of the combined complexity of nested relational algebra query evaluation. Elena Botoeva, Diego Calvanese, Benjamin Cogrel, Guohui Xiao 0001 |
ICDT | 2 |
| 2018 | Efficient Handling of SPARQL OPTIONAL for OBDA
Guohui Xiao 0001, Roman Kontchakov, Benjamin Cogrel, Diego Calvanese, Elena Botoeva |
ISWC (1) | 4 |
| 2018 | Semantics, Analysis and Simplification of DMN Decision Tables
Diego Calvanese, Marlon Dumas, Ülari Laurson, Fabrizio Maria Maggi, Marco Montali, Irene Teinemaa |
Inf. Syst. | 1 |
| 2017 | Cost-Driven Ontology-Based Data Access
Davide Lanti, Guohui Xiao 0001, Diego Calvanese |
ISWC (1) | 3 |
| 2016 | A semantic approach to polystoresabstractIn the database community Polystores is an emerging and promising approach for data federation that aims at designing a unified querying layer over multiple data models. In the Semantic Web community a similar in spirit approach of Ontology-Based Data Access (OBDA) has been recently proposed, attracted a lot of attention, and proved its success in several industrial scenarios. In this paper we discuss a semantic approach to building polystores using the OBDA paradigm. We also present our system Optique that is utilized in an industrial application of performing turbine diagnostics in Siemens. Evgeny Kharlamov, Theofilos P. Mailis, Konstantina Bereta, Dimitris Bilidas, Sebastian Brandt 0001, Ernesto Jiménez-Ruiz, Steffen Lamparter, Christian Neuenstadt, Özgür L. Özçep, Ahmet Soylu, Christoforos Svingos, Guohui Xiao 0001, Dmitriy Zheleznyakov, Diego Calvanese, Ian Horrocks 0001, Martin Giese, Yannis E. Ioannidis, Yannis Kotidis, Ralf Möller 0001, Arild Waaler |
IEEE BigData | 14 |
| 2016 | Handling Inconsistencies Due to Class Disjointness in SPARQL Updates
Albin Ahmeti, Diego Calvanese, Axel Polleres, Vadim Savenkov |
ESWC | 2 |
| 2016 | Verification of Evolving Graph-structured Data under Expressive Path ConstraintsabstractIntegrity constraints play a central role in databases and, among other applications, are fundamental for preserving data integrity when databases evolve as a result of operations manipulating the data. In this context, an important task is that of static verification, which consists in deciding whether a given set of constraints is preserved after the execution of a given sequence of operations, for every possible database satisfying the initial constraints. In this paper, we consider constraints over graph-structured data formulated in an expressive Description Logic (DL) that allows for regular expressions over binary relations and their inverses, generalizing many of the well-known path constraint languages proposed for semi-structured data in the last two decades. In this setting, we study the problem of static verification, for operations expressed in a simple yet flexible language built from additions and deletions of complex DL expressions. We establish undecidability of the general setting, and identify suitable restricted fragments for which we obtain tight complexity results, building on techniques developed in our previous work for simpler DLs. As a by-product, we obtain new (un)decidability results for the implication problem of path constraints, and improve previous upper bounds on the complexity of the problem. Diego Calvanese, Magdalena Ortiz 0001, Mantas Simkus |
ICDT | 1 |
| 2015 | The NPD Benchmark: Reality Check for OBDA SystemsabstractIn the last decades we moved from a world in which an enterprise had one central database---rather small for todays' standards---to a world in which many different---and big---databases must interact and operate, providing the user an integrated and understandable view of the data. Ontology-Based Data Access (OBDA) is becoming a popular approach to cope with this new scenario. OBDA separates the user from the data sources by means of a conceptual view of the data (ontology) that provides clients with a convenient query vocabulary. The ontology is connected to the data sources through a declarative specification given in terms of mappings. Although prototype OBDA systems providing the ability to answer SPARQL queries over the ontology are available, a significant challenge remains when it comes to use these systems in industrial environments: performance. To properly evaluate OBDA systems, benchmarks tailored towards the requirements in this setting are needed. In this work, we propose a novel benchmark for OBDA systems based on real data coming from the oil industry: the Norwegian Petroleum Directorate (NPD) FactPages. Our benchmark comes with novel techniques to generate, from the NPD data, datasets of increasing size, taking into account the requirements dictated by the OBDA setting. We validate our benchmark on significant OBDA systems, showing that it is more adequate than previous benchmarks not tailored for OBDA. Davide Lanti, Martín Rezk, Guohui Xiao 0001, Diego Calvanese |
EDBT | 4 |
| 2015 | Ontology-Based Integration of Cross-Linked Datasets
Diego Calvanese, Martin Giese, Dag Hovland, Martín Rezk |
ISWC (1) | 1 |
| 2015 | Special issue of the Journal of Web Semantics on ontology-based data access
Diego Calvanese, Manolis Koubarakis, David Toman 0001 |
J. Web Semant. | 1 |
| 2014 | Verifiable UML Artifact-Centric Business Process ModelsabstractArtifact-centric business process models have gained increasing momentum recently due to their ability to combine structural (i.e., data related) with dynamical (i.e., process related) aspects. In particular, two main lines of research have been pursued so far: one tailored to business artifact modeling languages and methodologies, the other focused on the foundations for their formal verification. In this paper, we merge these two lines of research, by showing how recent theoretical decidability results for verification can be fruitfully transferred to a concrete UML-based modeling methodology. In particular, we identify additional steps in the methodology that, in significant cases, guarantee the possibility of verifying the resulting models against rich first-order temporal properties. Notably, our results can be seamlessly transferred to different languages for the specification of the artifact lifecycles. Diego Calvanese, Marco Montali, Montserrat Estañol, Ernest Teniente |
CIKM | 1 |
| 2014 | Updating RDFS ABoxes and TBoxes in SPARQL
Albin Ahmeti, Diego Calvanese, Axel Polleres |
ISWC (1) | 2 |
| 2013 | Foundations of data-aware process analysis: a database theory perspectiveabstractIn this work we survey the research on foundations of data-aware (business) processes that has been carried out in the database theory community. We show that this community has indeed developed over the years a multi-faceted culture of merging data and processes. We argue that it is this community that should lay the foundations to solve, at least from the point of view of formal analysis, the dichotomy between data and processes still persisting in business process management. Diego Calvanese, Giuseppe De Giacomo, Marco Montali |
PODS | 1 |
| 2013 | Verification of relational data-centric dynamic systems with external servicesabstractData-centric dynamic systems are systems where both the process controlling the dynamics and the manipulation of data are equally central. We study verification of (first-order) mu-calculus variants over relational data-centric dynamic systems, where data are maintained in a relational database, and the process is described in terms of atomic actions that evolve the database. Action execution may involve calls to external services, thus inserting fresh data into the system. As a result such systems are infinite-state. We show that verification is undecidable in general, and we isolate notable cases where decidability is achieved. Specifically we start by considering service calls that return values deterministically (depending only on passed parameters). We show that in a mu-calculus variant that preserves knowledge of objects appeared along a run we get decidability under the assumption that the fresh data introduced along a run are bounded, though they might not be bounded in the overall system. In fact we tie such a result to a notion related to weak acyclicity studied in data exchange. Then, we move to nondeterministic services and we investigate decidability under the assumption that knowledge of objects is preserved only if they are continuously present. We show that if infinitely many values occur in a run but do not accumulate in the same state, then we get again decidability. We give syntactic conditions to avoid this accumulation through the novel notion of "generate-recall acyclicity", which ensures that every service call activation generates new values that cannot be accumulated indefinitely. Babak Bagheri Hariri, Diego Calvanese, Giuseppe De Giacomo, Alin Deutsch, Marco Montali |
PODS | 2 |
| 2012 | OCL-Lite: Finite reasoning on UML/OCL conceptual schemas
Anna Queralt, Alessandro Artale, Diego Calvanese, Ernest Teniente |
Data Knowl. Eng. | 3 |
| 2012 | Query Processing under GLAV Mappings for Relational and Graph DatabasesabstractSchema mappings establish a correspondence between data stored in two databases, called source and target respectively. Query processing under schema mappings has been investigated extensively in the two cases where each target atom is mapped to a query over the source (called GAV, global-as-view), and where each source atom is mapped to a query over the target (called LAV, local-as-view). The general case, called GLAV, in which queries over the source are mapped to queries over the target, has attracted a lot of attention recently, especially for data exchange. However, query processing for GLAV mappings has been considered only for the basic service of query answering, and mainly in the context of conjunctive queries (CQs) in relational databases. In this paper we study query processing for GLAV mappings in a wider sense, considering not only query answering, but also query rewriting, perfectness (the property of a rewriting to compute exactly the certain answers), and query containment relative to a mapping. We deal both with the relational case, and with graph databases, where the basic querying mechanism is that of regular path queries. Query answering in GLAV can be smoothly reduced to a combination of the LAV and GAV cases, and for CQs this reduction can be exploited also for the remaining query processing tasks. In contrast, as we show, GLAV query processing for graph databases is non-trivial and requires new insights and techniques. We obtain upper bounds for answering, rewriting, and perfectness, and show decidability of relative containment. Diego Calvanese, Giuseppe De Giacomo, Maurizio Lenzerini, Moshe Y. Vardi |
Proc. VLDB Endow. | 1 |
| 2011 | Simplifying schema mappingsabstractA schema mapping is a formal specification of the relationship holding between the databases conforming to two given schemas, called source and target, respectively. While in the general case a schema mapping is specified in terms of assertions relating two queries in some given language, various simplified forms of mappings, in particular LAV and GAV, have been considered, based on desirable properties that these forms enjoy. Recent works propose methods for transforming schema mappings to logically equivalent ones of a simplified form. In many cases, this transformation is impossible, and one might be interested in finding simplifications based on a weaker notion, namely logical implication, rather than equivalence. More precisely, given a schema mapping M, find a simplified (LAV, or GAV) schema mapping M' such that M' logically implies M. In this paper we formally introduce this problem, and study it in a variety of cases, providing techniques and complexity bounds. The various cases we consider depend on three parameters: the simplified form to achieve (LAV, or GAV), the type of schema mapping considered (sound, or exact), and the query language used in the schema mapping specification (conjunctive queries and variants over relational databases, or regular path queries and variants over graph databases). Notably, this is the first work on comparing schema mappings for graph databases. Diego Calvanese, Giuseppe De Giacomo, Maurizio Lenzerini, Moshe Y. Vardi |
ICDT | 1 |
| 2010 | Full Satisfiability of UML Class Diagrams
Alessandro Artale, Diego Calvanese, Yazmín Ibáñez-García |
ER | 2 |
| 2010 | Evolution of DL-Lite Knowledge Bases
Diego Calvanese, Evgeny Kharlamov, Werner Nutt, Dmitriy Zheleznyakov |
ISWC (1) | 1 |
| 2009 | Discovering functional dependencies for multidimensional designabstractNowadays, it is widely accepted that the data warehouse design task should be largely automated. Furthermore, the data warehouse conceptual schema must be structured according to the multidimensional model and as a consequence, the most common way to automatically look for subjects and dimensions of analysis is by discovering functional dependencies (as dimensions functionally depend on the fact) over the data sources. Most advanced methods for automating the design of the data warehouse carry out this process from relational OLTP systems, assuming that a RDBMS is the most common kind of data source we may find, and taking as starting point a relational schema. In contrast, in our approach we propose to rely instead on a conceptual representation of the domain of interest formalized through a domain ontology expressed in the DL-Lite Description Logic. We propose an algorithm to discover functional dependencies from the domain ontology that exploits the inference capabilities of DL-Lite, thus fully taking into account the semantics of the domain. We also provide an evaluation of our approach in a real-world scenario. Oscar Romero 0001, Diego Calvanese, Alberto Abelló, Mariano Rodriguez-Muro |
DOLAP | 2 |
| 2009 | Controlled Aggregate Tree Shaped Questions over Ontologies
Camilo Thorne, Diego Calvanese |
FQAS | 2 |
| 2008 | Inconsistency tolerance in P2P data integration: An epistemic logic approach
Diego Calvanese, Giuseppe De Giacomo, Domenico Lembo, Maurizio Lenzerini, Riccardo Rosati 0001 |
Inf. Syst. | 1 |
| 2007 | Reasoning over Extended ER Models
Alessandro Artale, Diego Calvanese, Roman Kontchakov, Vladislav Ryzhikov, Michael Zakharyaschev |
ER | 2 |
| 2006 | Enterprise modeling and Data Warehousing in Telecom Italia
Diego Calvanese, Luigi Dragone, Daniele Nardi, Riccardo Rosati 0001, Stefano Trisolini |
Inf. Syst. | 1 |
| 2005 | View-Based Query Processing: On the Relationship Between Rewriting, Answering and Losslessness
Diego Calvanese, Giuseppe De Giacomo, Maurizio Lenzerini, Moshe Y. Vardi |
ICDT | 1 |
| 2005 | Automatic Composition of Transition-based Semantic Web Services with Messaging
Daniela Berardi, Diego Calvanese, Giuseppe De Giacomo, Richard Hull 0001, Massimo Mecella |
VLDB | 2 |
| 2005 | Automatic Service Composition Based on Behavioral DescriptionsabstractThis paper addresses the issue of automatic service composition. We first develop a framework in which the exported behavior of a service is described in terms of a so-called execution tree, that is an abstraction for its possible executions. We then study the case in which such exported behavior (i.e. the execution tree of the service) can be represented by a finite state machine (i.e. finite state transition system). In this specific setting, we devise sound, complete and terminating techniques both to check for the existence of a composition, and to return a composition, if one exists. We also analyze the computational complexity of the proposed algorithms. Finally, we present an open source prototype tool, called [Formula: see text] (E-Service Composer), that implements our composition technique. To the best of our knowledge, our work is the first attempt to provide a provably correct technique for the automatic synthesis of service composition, in a framework where the behavior of services is explicitly specified. Daniela Berardi, Diego Calvanese, Giuseppe De Giacomo, Maurizio Lenzerini, Massimo Mecella |
Int. J. Cooperative Inf. Syst. | 2 |
| 2004 | Logical Foundations of Peer-To-Peer Data IntegrationabstractIn peer-to-peer data integration, each peer exports data in terms of its own schema, and data interoperation is achieved by means of mappings among the peer schemas. Peers are autonomous systems and mappings are dynamically created and changed. One of the challenges in these systems is answering queries posed to one peer taking into account the mappings. Obviously, query answering strongly depends on the semantics of the overall system. In this paper, we compare the commonly adopted approach of interpreting peerto-peer systems using a first-order semantics, with an alternative approach based on epistemic logic. We consider several central properties of peer-to-peer systems: modularity, generality, and decidability. We argue that the approach based on epistemic logic is superior with respect to all the above properties. In particular, we show that, in systems in which peers have decidable schemas and conjunctive mappings, but are arbitrarily interconnected, the first-order approach may lead to undecidability of query answering, while the epistemic approach always preserves decidability. This is a fundamental property, since the actual interconnections among peers are not under the control of any actor in the system. 1. Diego Calvanese, Giuseppe De Giacomo, Maurizio Lenzerini, Riccardo Rosati 0001 |
PODS | 1 |
| 2004 | Data integration under integrity constraints
Andrea Calì, Diego Calvanese, Giuseppe De Giacomo, Maurizio Lenzerini |
Inf. Syst. | 2 |
| 2003 | IBIS: Semantic Data Integration at Work
Andrea Calì, Diego Calvanese, Giuseppe De Giacomo, Maurizio Lenzerini, Paolo Naggar, Fabio Vernacotola |
CAiSE | 2 |
| 2003 | Decidable Containment of Recursive Queries
Diego Calvanese, Giuseppe De Giacomo, Moshe Y. Vardi |
ICDT | 1 |
| 2003 | View-based query containmentabstractQuery containment is the problem of checking whether for all databases the answer to a query is a subset of the answer to a second query. In several data management tasks, such as data integration, mobile computing, etc., the data of interest are only accessible through a given set of views. In this case, containment of queries should be determined relative to the set of views, as already noted in the literature. Such a form of containment, which we call view-based query containment, is the subject of this paper. The problem comes in various forms, depending on whether each of the two queries is expressed over the base alphabet or the alphabet of the view names. We present a thorough analysis of view-based query containment, by discussing all possible combinations from a semantic point of view, and by showing their mutual relationships. In particular, for the two settings of conjunctive queries and two-way regular path queries, we provide both techniques and complexity bounds for the different variants of the problem. Finally, we study the relationship between view-based query containment and view-based query rewriting. Diego Calvanese, Giuseppe De Giacomo, Maurizio Lenzerini, Moshe Y. Vardi |
PODS | 1 |
| 2002 | Data Integration under Integrity Constraints
Andrea Calì, Diego Calvanese, Giuseppe De Giacomo, Maurizio Lenzerini |
CAiSE | 2 |
| 2002 | On the Expressive Power of Data Integration Systems
Andrea Calì, Diego Calvanese, Giuseppe De Giacomo, Maurizio Lenzerini |
ER | 2 |
| 2002 | Lossless Regular ViewsabstractIf the only information we have on a certain database is through a set of views, the question arises of whether this is sufficient to answer completely a given query. We say that the set of views is lossless with respect to the query, if, no matter what the database is, we can answer the query by solely relying on the content of the views. The question of losslessness has various applications, for example in query optimization, mobile computing, data warehousing, and data integration. We study this problem in a context where the database is semistructured, and both the query and the views are expressed as regular path queries. The form of recursion present in this class prevents us from applying known results to our case.We first address the problem of checking losslessness in the case where the views are materialized. The fact that we have the view extensions available makes this case solvable by extending known techniques. We then study a more complex version of the problem, namely the one where we abstract from the specific view extension. More precisely, we address the problem of checking whether, for every database, the answer to the query over such a database can be obtained by relying only on the view extensions. We show that the problem is solvable by utilizing, via automata-theoretic techniques, the known connection between view-based query answering and constraint satisfaction. We also investigate the computational complexity of both versions of the problem. Diego Calvanese, Giuseppe De Giacomo, Maurizio Lenzerini, Moshe Y. Vardi |
PODS | 1 |
| 2001 | Accessing Data Integration Systems through Conceptual Schemas
Andrea Calì, Diego Calvanese, Giuseppe De Giacomo, Maurizio Lenzerini |
ER | 2 |
| 2001 | Data Integration in Data WarehousingabstractInformation integration is one of the most important aspects of a Data Warehouse. When data passes from the sources of the application-oriented operational environment to the Data Warehouse, possible inconsistencies and redundancies should be resolved, so that the warehouse is able to provide an integrated and reconciled view of data of the organization. We describe a novel approach to data integration in Data Warehousing. Our approach is based on a conceptual representation of the Data Warehouse application domain, and follows the so-called local-as-view paradigm: both source and Data Warehouse relations are defined as views over the conceptual model. We propose a technique for declaratively specifying suitable reconciliation correspondences to be used in order to solve conflicts among data in different sources. The main goal of the method is to support the design of mediators that materialize the data in the Data Warehouse relations. Starting from the specification of one such relation as a query over the conceptual model, a rewriting algorithm reformulates the query in terms of both the source relations and the reconciliation correspondences, thus obtaining a correct specification of how to load the data in the materialized view. Diego Calvanese, Giuseppe De Giacomo, Maurizio Lenzerini, Daniele Nardi, Riccardo Rosati 0001 |
Int. J. Cooperative Inf. Syst. | 1 |
| 2000 | Answering Regular Path Queries Using ViewsabstractQuery answering using views amounts to computing the answer to a query having information only on the extension of a set of views. This problem is relevant in several fields, such as information integration, data warehousing, query optimization, mobile computing, and maintaining physical data independence. We address query answering using views in a context where queries and views are regular path queries, i.e., regular expressions that denote the pairs of objects in the database connected by a matching path. Regular path queries are the basic query mechanism when the database is conceived as a graph, such as in semistructured data and data on the Web. We study algorithms for answering regular path queries using views under different assumptions, namely, closed and open domain, and sound, complete, and exact information on view extensions. We characterize data, expression, and combined complexity of the problem, showing that the proposed algorithms are essentially optimal. Our results are the first to exhibit decidability in cases where the language for expressing the query and the views allows for recursion. Diego Calvanese, Giuseppe De Giacomo, Maurizio Lenzerini, Moshe Y. Vardi |
ICDE | 1 |
| 2000 | View-Based Query Processing for Regular Path Queries with InverseabstractView-based query processing is the problem of computing the answer to a query based on a set of materialized views, rather than on the raw data in the database. The problem comes in two different forms, called query rewriting and query answering, respectively. In the first form, we are given a query and a set of view definitions, and the goal is to reformulate the query into an expression that refers only to the views. In the second form, besides the query and the view definitions, we are also given the extensions of the views and a tuple, and the goal is to check whether the knowledge on the view extensions logically implies that the tuple satisfies the query. In this paper we address the problem of view-based query processing in the context of semistructured data, in particular for the case of regular-path queries extended with the inverse operator. Several authors point out that the inverse operator is one of the fundamental extensions for making regular-path queries useful in real settings. We present a novel technique based on the use of two-way finite-state automata. Our approach demonstrates the power of this kind of automata in dealing with the inverse operator, allowing us to show that both query rewriting and query answering with the inverse operator has the same computational complexity as for the case of standard regular-path queries. 1. Diego Calvanese, Giuseppe De Giacomo, Maurizio Lenzerini, Moshe Y. Vardi |
PODS | 1 |
| 2000 | Concept Based Design of Data Warehouses: The DWQ DemonstratorsabstractThe ESPRIT Project DWQ (Foundations of Data Warehouse Quality) aimed at improving the quality of DW design and operation through systematic enrichment of the semantic foundations of data warehousing. Logic-based knowledge representation and reasoning techniques were developed to control accuracy, consistency, and completeness via advanced conceptual modeling techniques for source integration, data reconciliation, and multi-dimensional aggregation. This is complemented by quantitative optimization techniques for view materialization, optimizing timeliness and responsiveness without losing the semantic advantages from the conceptual approach. At the operational level, query rewriting and materialization refreshment algorithms exploit the knowledge developed at design time. The demonstration shows the interplay of these tools under a shared metadata repository, based on an example extracted from an application at Telecom Italia. Matthias Jarke, Christoph Quix, Diego Calvanese, Maurizio Lenzerini, Enrico Franconi, Spyros Ligoudistianos, Panos Vassiliadis, Yannis Vassiliou |
SIGMOD Conference | 3 |
| 1999 | Queries and Constraints on Semi-structured Data
Diego Calvanese, Giuseppe De Giacomo, Maurizio Lenzerini |
CAiSE | 1 |
| 1999 | Rewriting of Regular Expressions and Regular Path QueriesabstractRecent work on semi-structured da.ta ha.s revitalized the interest in pa.th qu.eries, i.e., queries that ask for ah pairs of objects in the database that are connected by a, path conforming to a certain specification, in particular to a regular expression.Also, in semi-structured data., as well as in data.integration, da.ta.wa.rehousing, and query optimization, the problem of query rewriting using views is receiving much attention: Given a. query and a collection of views, generate a new query which uses the views and provides the answer to the original one.In this paper we address the problem of query rewriting using views in the context of semi-structured data.We present a method for computing the rewriting of a regular expression i? in terms of other regular expressions.The method computes the exact rewriting (the one that defines the same regular language as E) if it exists, or the rewriting that defines the maximal language contained in the one defined by E, otherwise.We present a complexity analysis of both the problem+and the method, showing that the latter is essentially optimal.Finally, we illustrate how to exploit the method to rewrite regular path queries using views in semistructured data.The complexity results established for the rewriting of regular expressions apply also to the case of regu1a.rpath queries. Diego Calvanese, Giuseppe De Giacomo, Maurizio Lenzerini, Moshe Y. Vardi |
PODS | 1 |
| 1998 | On the Decidability of Query Containment under ConstraintsabstractQuery containment under constraints is the problem of checking whether for every database satisfying a given set of constraints, the result of one query is a subset of the result of another query, Recent research points out that this is a central problem in severa database applications, and we address it within A setting where constraints are specified in the form of special inclusion dependencies over complex expressions, built by using intersection and difference of relations, special forms of quantification, regular expressions over binary relations, and cardinality constraints.These types of constraints capture a great variety of data models, including the relational, the entity-relational, and the object-oriented model,We study the problem of checking whether q is contained in q' with respect to the constraints specified in a schema S, where q and q' are nonrecursive Datalog programs whose atoms are complex expressions.We present the following results on query containment.For the case where q does not contain regular expressions, we provide a method for deciding query containment, and analyze its computational complexity.We do the same for the case where neither S nor q, q' contain number restrictions.To the best of our knowledge, this yields the first decidability result on containment of conjunctive queries with regular expressions.Finally, we Provo that the problem is undecidable for the case where we admit inequalities in q', , 1 Introduction Diego Calvanese, Giuseppe De Giacomo, Maurizio Lenzerini |
PODS | 1 |
| 1994 | On the Interaction Between ISA and Cardinality ConstraintsabstractISA and cardinality constraints are among the most interesting types of constraints in data models. ISA constraints are used to establish several forms of containment among classes, and are receiving great attention in moving to object-oriented data models, where classes are organized in hierarchies based on a generalization/specialization principle. Cardinality constraints impose restrictions on the number of links of a certain type involving every instance of a given class, and can be used for representing several forms of dependencies between classes, including functional and existence dependencies. While the formal properties of each type of constraints are now well understood, little is known of their interaction. We present an effective method for reasoning about a set of ISA and cardinality constraints in the context of a simple data model based on the notions of classes and relationships. In particular, the method allows one both to verify the satisfiability of a schema and to check whether a schema implies a given constraint of any of the two kinds. We prove that the method is sound and complete, thus showing that the reasoning problem for ISA and cardinality constraints is decidable.> Diego Calvanese, Maurizio Lenzerini |
ICDE | 1 |
| 1994 | Making Object-Oriented Schemas More ExpressiveabstractCurrent object-oriented data models lack several important features that would allow one to express relevant knowledge about the classes of schema. In particular, there is no data model supporting simultaneously the inverse of the functions represented by attributes, the union, the intersection and the complement of classes, the possibility of using nonbinary relations, and the possibility of expressing cardinality constraints on attributes and relations. In this paper we define a new data model, called CAR, which extends the basic core of current object-oriented data models with all the above mentioned features. A technique is then presented both for checking the consistency of class definitions, and for computing the logical sequences of the knowledge represented in the schema. Finally, the inherent complexity of reasoning in CAR is investigated, and the complexity of our inferencing technique is studied, depending on various assumptions on the schema. Diego Calvanese, Maurizio Lenzerini |
PODS | 1 |