Andreas K. Maier

dblp:131/7133 · also Andreas Maier 0001 · DBLP profile ↗
← Back
13ranked-venue papers in the field
0as first author
5since 2021 · last 2026
0000-0002-9550-5284ORCID · conflict

Domains — venue-derived; a paper can count in several

Other / Interdisciplinary · 13
YearPublicationVenuePosition
2026 From Pixels to Structure: Lightweight Vision-Language Models for Document OCR and Structured JSON Extraction
Uddipan Basu Bir, Vincent Christlein, Andreas K. Maier, Mathias Zinnen
ICDAR (3)3
2022 Is Multitask Learning Always Better?
Alexander Mattick, Martin Mayr, Andreas K. Maier, Vincent Christlein
DAS3
2022 Combining Visual and Linguistic Models for a Robust Recipient Line Recognition in Historical Documents
Martin Mayr, Alex Felker, Andreas K. Maier, Vincent Christlein
DAS3
2021 SmartPatch: Improving Handwritten Word Imitation with Patch Discriminators
Alexander Mattick, Martin Mayr, Mathias Seuret, Andreas K. Maier, Vincent Christlein
ICDAR (1)4
2021 ICDAR 2021 Competition on Historical Document Classification
Mathias Seuret, Anguelos Nicolaou, Dalia Rodríguez-Salas, Nikolaus Weichselbaumer, Dominique Stutzmann, Martin Mayr, Andreas K. Maier, Vincent Christlein
ICDAR (4)7
2020 Re-Ranking for Writer Identification and Writer Retrieval
Simon Jordan, Mathias Seuret, Pavel Král, Ladislav Lenc, Jirí Martínek, Barbara Wiermann, Tobias Schwinger, Andreas K. Maier, Vincent Christlein
DAS8
2020 The Notary in the Haystack - Countering Class Imbalance in Document Processing with CNNs
Martin Leipert, Georg Vogeler, Mathias Seuret, Andreas K. Maier, Vincent Christlein
DAS4
2019 ICDAR 2019 Competition on Image Retrieval for Historical Handwritten Documents
abstract
This competition investigates the performance of large-scale retrieval of historical document images based on writing style. Based on large image data sets provided by cultural heritage institutions and digital libraries, providing a total of 20 000 document images representing about 10 000 writers, divided in three types: writers of (i) manuscript books, (ii) letters, (iii) charters and legal documents. We focus on the task of automatic image retrieval to simulate common scenarios of humanities research, such as writer retrieval. The most teams submitted traditional methods not using deep learning techniques. The competition results show that a combination of methods is outperforming single methods. Furthermore, letters are much more difficult to retrieve than manuscripts.
Vincent Christlein, Anguelos Nicolaou, Mathias Seuret, Dominique Stutzmann, Andreas K. Maier
ICDAR5
2019 Deep Generalized Max Pooling
abstract
Global pooling layers are an essential part of Convolutional Neural Networks (CNN). They are used to aggregate activations of spatial locations to produce a fixed-size vector in several state-of-the-art CNNs. Global average pooling or global max pooling are commonly used for converting convolutional features of variable size images to a fix-sized embedding. However, both pooling layer types are computed spatially independent: each individual activation map is pooled and thus activations of different locations are pooled together. In contrast, we propose Deep Generalized Max Pooling that balances the contribution of all activations of a spatially coherent region by re-weighting all descriptors so that the impact of frequent and rare ones is equalized. We show that this layer is superior to both average and max pooling on the classification of Latin medieval manuscripts (CLAMM'16, CLAMM'17), as well as writer identification (Historical-WI'17).
Vincent Christlein, Lukas Spranger, Mathias Seuret, Anguelos Nicolaou, Pavel Král, Andreas K. Maier
ICDAR6
2018 Encoding CNN Activations for Writer Recognition
abstract
The encoding of local features is an essential part for writer identification and writer retrieval. While CNN activations have already been used as local features in related works, the encoding of these features has attracted little attention so far. In this work, we compare the established VLAD encoding with triangulation embedding. We further investigate generalized max pooling as an alternative to sum pooling and the impact of decorrelation and Exemplar SVMs. With these techniques, we set new standards on two publicly available datasets (ICDAR13, KHATT).
Vincent Christlein, Andreas K. Maier
DAS2
2018 Non-destructive Digitization of Soiled Historical Chinese Bamboo Scrolls
abstract
For about 2000 years, no paper was used as a media in China but writings and drawings were captured on bamboo and wooden slips. Several slips were bound together with strips and rolled up to a scroll. The writings and drawings were either brushed or even carved into the wood. Those documents are very precious for culture inheritance and research, but due to aging processes, the discovered pieces are sometimes in a poor condition and also soiled. Because cleaning the slips is not only challenging but also writings could be erased, we developed a method to digitize such historical documents without the need of cleaning. We perform a 3-D X-ray micro-CT scan resulting in a 3-D volume of the complete document. With our approach, we were able to investigate the scroll without any manual labor (e.g. unwrapping or cleaning). We showed that the method also works for heavily soiled scrolls where nothing is readable with the naked eye. This can help conservators to store all writings before they may be erased by the cleaning process. Finally, we present a manual technique to virtually unwrap and post-process the documents resulting in a 2-D image of all bamboo slips.
Daniel Stromer, Vincent Christlein, Andreas K. Maier, Patrick Zippert, Eric Helmecke, Tino Hausotte, Xiaolin Huang
DAS3
2017 Unsupervised Feature Learning for Writer Identification and Writer Retrieval
abstract
Deep Convolutional Neural Networks (CNN) have shown great success in supervised classification tasks such as character classification or dating. Deep learning methods typically need a lot of annotated training data, which is not available in many scenarios. In these cases, traditional methods are often better than or equivalent to deep learning methods. In this paper, we propose a simple, yet effective, way to learn CNN activation features in an unsupervised manner. Therefore, we train a deep residual network using surrogate classes. The surrogate classes are created by clustering the training dataset, where each cluster index represents one surrogate class. The activations from the penultimate CNN layer serve as features for subsequent classification tasks. We evaluate the feature representations on two publicly available datasets. The focus lies on the ICDAR17 competition dataset on historical document writer identification (Historical-WI). We show that the activation features trained without supervision are superior to descriptors of state-of-the-art writer identification methods. Additionally, we achieve comparable results in the case of handwriting classification using the ICFHR16 competition dataset on historical Latin script types (CLaMM16).
Vincent Christlein, Martin Gropp, Stefan Fiel, Andreas K. Maier
ICDAR4
2017 Browsing through Closed Books: Fully Automatic Book Page Extraction from a 3-D X-Ray CT Volume
abstract
When digitizing or investigating historical documents, it is often the case that a document can not be opened, page-turned or touched anymore. Damages such as moisture or fire and aging processes disallow browsing through a book. To address these particular cases, our earlier work showed that Micro-CT X-ray scanners are able to image documents written with iron gall ink. A self-made book consisting of ten hand written pages was scanned and investigated without opening or page-turning. However, when analyzing the reconstruction results, we faced the problem of a proper automatic page segmentation and 2-D mapping within the volume in an acceptable time without losing information of the writings. The main problem is that the pages can be arbitrary deformed or squeezed together. In this paper, we present a fully automatic algorithm for the segmentation and extraction of book pages from the original 3-D volume. Our method delivers high quality results for our book model and can be easily adapted to other imaging modalities. We show that it performs well even for an extreme case with low resolution input data and wavy pages. To keep it simple for users, our algorithm works without any need of prior information or user interactions.
Daniel Stromer, Vincent Christlein, Tobias Schön, Wolfgang Holub, Andreas K. Maier
ICDAR5