Alexander Graß

dblp:194/7535 · also Alexander Grass · DBLP profile ↗
← Back
9ranked-venue papers
4as first author
6since 2021 · last 2026
0009-0002-1416-1676ORCID · corroborated

Domains — the database's venue-derived domains; a paper can count in several

Databases, data management, data science and information retrieval · 6 · 2 first-author · 3 since 2021Applied, interdisciplinary, general and emerging computing · 6 · 3 first-author · 3 since 2021Artificial intelligence and machine learning · 5 · 1 first-author · 2 since 2021Systems, architecture and hardware · 1 · 1 since 2021Theory of computation · 1 · 1 since 2021
YearPublicationVenuePosition
2026 SemTS: Ontology and Vocabularies for the Semantic Categorization of Time Series Knowledge
Alexander Graß, Rohit A. Deshmukh, Christoph Lange 0002, Diego Collarana, Christian Beecks, Stefan Decker
ESWC (2)1
2025 Code2Onto: Multi-Agent System for Code-Driven Ontology Population
abstract
5642
Alexander Graß, Jonathan Lehmkuhl, Diego Collarana, Stefan Decker, Christian Beecks
IEEE Big Data1
2025 Semantic Intelligence: Graph RAG-Driven Agents for Time Series Analytics
Alexander Graß, Christopher I. Pack, Diego Collarana, Stefan Decker, Christian Beecks
IDEAL (1)1
2024 Structural and Semantic Data Layers in Time Series Analyses
Alexander Graß, Christian Beecks, Stefan Decker
IDEAL (1)1
2023 Interpreting Black-box Machine Learning Models for High Dimensional Datasets
abstract
Many datasets are of increasingly high dimension- ality, where a large number of features could be irrelevant to the learning task. The inclusion of such features would not only introduce unwanted noise but also increase computational complexity. Deep neural networks (DNNs) outperform machine learning (ML) algorithms in a variety of applications due to their effectiveness in modelling complex problems and handling high-dimensional datasets. However, due to non-linearity and higher-order feature interactions, DNN models are unavoidably opaque, making them black-box methods. In contrast, an interpretable model can identify statistically significant features and explain the way they affect the model’s outcome. In this paper, we propose a novel method to improve the interpretability of blackbox models in the case of high-dimensional datasets. First, a black-box model is trained on full feature space that learns useful embeddings on which the classification is performed. To decompose the inner principles of the black-box and to identify top-k important features (global explainability), probing and perturbing techniques are applied. An interpretable surrogate model is then trained on top-k feature space to approximate the black-box. Finally, decision rules and counterfactuals are derived from the surrogate to provide local decisions. Our approach outperforms tabular learners, e.g., TabNet and XGboost, and SHAP-based interpretability techniques, when tested on a number of datasets having dimensionality between 54 and 20,5311.1GitHub: https://github.com/rezacsedu/DeepExplainHidim
Md. Rezaul Karim 0001, Md Shajalal, Alexander Graß, Till Döhmen, Sisay Adugna Chala, Alexander Boden, Christian Beecks, Stefan Decker
DSAA3
2021 knowlEdge Project -Concept, Methodology and Innovations for Artificial Intelligence in Industry 4.0
abstract
AI is one of the biggest megatrends towards the 4th industrial revolution. Although these technologies promise business sustainability as well as product and process quality, it seems that the ever-changing market demands, the complexity of technologies and fair concerns about privacy, impede broad application and reuse of Artificial Intelligence (AI) models across the industry. To break the entry barriers for these technologies and unleash its full potential, the knowlEdge project will develop a new generation of AI methods, systems, and data management infrastructure. Subsequently, as part of the knowlEdge project we propose several major innovations in the areas of data management, data analytics and knowledge management including (i) a set of AI services that allows the usage of edge deployments as computational and live data infrastructure as well as a continuous learning execution pipeline on the edge, (ii) a digital twin of the shop-floor able to test AI models, (iii) a data management framework deployed along the edge-to-cloud continuum ensuring data quality, privacy and confidentiality, (iv) Human-AI Collaboration and Domain Knowledge Fusion tools for domain experts to inject their experience into the system, (v) a set of standardisation mechanisms for the exchange of trained AI models from one context to another, and (vi) a knowledge marketplace platform to distribute and interchange trained AI models. In this paper, we present a short overview of the EU Project knowlEdge –Towards Artificial Intelligence powered manufacturing services, processes, and products in an edge-to-cloud-knowledge continuum for humans [in-the-loop], which is funded by the Horizon 2020 (H2020) Framework Programme of the European Commission under Grant Agreement 957331. Our overview includes a description of the project’s main concept and methodology as well as the envisioned innovations.
Sergio Álvarez-Napagao, Boki Ashmore, Marta Barroso, Cristian Barrué, Christian Beecks, Fabian Berns, Ilaria Bosi, Sisay Adugna Chala, Nicola Ciulli, Marta Garcia-Gasulla, Alexander Graß, Dimosthenis Ioannidis, Natalia Jakubiak, Karl Köpke, Ville Lämsä, Pedro Megias, Alexandros Nizamis, Claudio Pastrone, Rosaria Rossini, Miquel Sànchez-Marrè, Luca Ziliotti
INDIN11
2019 A New Approach for Efficient Structure Discovery in IoT
abstract
Complex, multivariate data streams frequently comprise subjacent behavioral patterns, which are subsumable by a process of statistical structure discovery. Revealing these hidden patterns from raw data is a major challenge in abstracting information and thus for new opportunities of efficient data analysis at scale. State-of-the-art approaches, such as CKS and ABCD, leverage statistical data models and Gaussian Processes in order to abstract from raw data and to describe their major data characteristics by means of kernel-decomposed covariance functions. The process of identifying the most appropriate covariance function is a performance bottleneck due to its super-quadratic computation time complexity for model selection and evaluation. In this paper, we thus propose a new approach for the computation of large-scale statistical data models. To this end, we propose to bound the complexity of the statistical data model and develop a sequential agglomerative approach to reduce the computational load of the required evaluative calculations. Our performance analysis indicates that our proposal is able to outperform state-of-the-art kernel search algorithms such as CKS and ABCD with respect to the qualities of efficiency and accuracy.
Fabian Berns, Kjeld Schmidt, Alexander Graß, Christian Beecks
IEEE BigData3
2018 Metric Indexing for Efficient Data Access in the Internet of Things
abstract
Data are a central phenomenon in our digital information age. They impact the way we live, work, and play and provide unprecedented opportunities to simplify our daily life and behavior. They implicate enormous potential and impact society, economy, and science. Due to the advancement of cyber-physical systems and Internet of Things technologies, it is expected that the majority of real-time data will be generated from devices interconnected within the Internet of Things by the year 2025. In this paper, we tackle the problem of managing Internet of Things data in an efficient way. To this end, we introduce the metric approach for storing and querying Internet of Things data and investigate the ability of pivot-based tables for indexing and searching this type of data. Along with the introduction of two real-world, large-scale Internet of Things datasets from the EU projects COMPOSITION and MONSOON (under grant no. 723145 and 723650), we show that the metric approach facilitates efficient data access in the Internet of Things.
Christian Beecks, Alexander Graß, Shreekantha Devasya
IEEE BigData2
2016 Multi-step threshold algorithm for efficient feature-based query processing in large-scale multimedia databases
abstract
Accessing very large multimedia databases in a content-based way has become one of the major challenges in todays' multimedia analysis and retrieval applications. Accompanied by the heterogeneity of data and the continuous change of user requirements, content-based approaches are supposed to efficiently retrieve and analyze query-like multimedia objects with the highest possible degree of efficacy, which makes the utilization of complex multimedia object representations and adaptive similarity measures inevitable. In order to facilitate efficient and flexible content-based access into multimedia databases comprising millions of complex data objects, we propose the Multi-step Threshold Algorithm (MTA). Based on a set of query features, the MTA aims to retrieve the most similar multimedia objects with minimal I/O cost by incrementally traversing an in-memory index structure in a feature-by-feature manner in order to approximate the multimedia objects' similarities prior to database access. In addition to the MTA, we propose different enhancements that ensure scalable feature-based query processing. Our performance analysis evidences that our proposal is able to process feature-based queries on a million-scale multimedia database in milliseconds on a single CPU.
Christian Beecks, Alexander Graß
IEEE BigData2