EDBT 2026 Demo / reviewers in the wild / expert
Vincent Poulain D'Andecy
dblp:125/8113
· DBLP profile ↗
20ranked-venue papers in the field
3as first author
8since 2021 · last 2025
—ORCID · unresolved
Domains — venue-derived; a paper can count in several
Other / Interdisciplinary · 20 (3 first)
| Year | Publication | Venue | Position |
|---|---|---|---|
| 2025 | Few-Shot Document Classification in Real Applications: Boosting Precision with Novelty Detection
Tri-Cong Pham, Mickaël Coustaty, Aurélie Joseph, Gaspar Deloin, Vincent Poulain D'Andecy, Antoine Doucet |
ICDAR (3) | 5 |
| 2025 | QUEST: Quality-Aware Semi-supervised Table Extraction for Business Documents
Eliott Thomas, Mickaël Coustaty, Aurélie Joseph, Gaspar Deloin, Elodie Carel, Vincent Poulain D'Andecy, Jean-Marc Ogier |
ICDAR (5) | 6 |
| 2024 | CHIC: Corporate Document for Visual Question Answering
Ibrahim Souleiman Mahamoud, Mickaël Coustaty, Aurélie Joseph, Vincent Poulain D'Andecy, Jean-Marc Ogier |
ICDAR (6) | 4 |
| 2024 | Privacy-Aware Document Visual Question Answering
Rubèn Tito, Marlon Tobaben, Raouf Kerkouche, Mohamed Ali Souibgui, Kangsoo Jung, Joonas Jälkö, Vincent Poulain D'Andecy, Aurélie Joseph, Lei Kang 0002, Ernest Valveny, Antti Honkela, Mario Fritz, Dimosthenis Karatzas |
ICDAR (6) | 8 |
| 2023 | Incremental Learning and Ambiguity Rejection for Document Classification
Tri-Cong Pham, Mickaël Coustaty, Aurélie Joseph, Vincent Poulain D'Andecy, Muriel Visani, Nicolas Sidere |
ICDAR (5) | 4 |
| 2023 | Receipt Dataset for Document Forgery Detection
Beatriz Martínez Tornés, Théo Taburet, Emanuela Boros, Kais Rouis, Antoine Doucet, Petra Gomez-Krämer, Nicolas Sidere, Vincent Poulain D'Andecy |
ICDAR (3) | 8 |
| 2022 | QAlayout: Question Answering Layout Based on Multimodal Attention for Visual Question Answering on Corporate Document
Ibrahim Souleiman Mahamoud, Mickaël Coustaty, Aurélie Joseph, Vincent Poulain D'Andecy, Jean-Marc Ogier |
DAS | 4 |
| 2021 | Multimodal Attention-Based Learning for Imbalanced Corporate Documents Classification
Ibrahim Souleiman Mahamoud, Joris Voerman, Mickaël Coustaty, Aurélie Joseph, Vincent Poulain D'Andecy, Jean-Marc Ogier |
ICDAR (3) | 5 |
| 2020 | Evaluation of Neural Network Classification Systems on Document Stream
Joris Voerman, Aurélie Joseph, Mickaël Coustaty, Vincent Poulain D'Andecy, Jean-Marc Ogier |
DAS | 4 |
| 2019 | Discourse Descriptor for Document Incremental Classification Comparison with Deep LearningabstractWe propose both a new strategy to weight a text vector for document classification and a comparison of a deep learning approach versus an incremental classification approach, integrating our novel strategy. Bag-of-word vectors are classic approaches to describe a textual document in document classification objective. A weakness of the bag of words is to lose the organization of the discourse within the document. Inspired by some Deep Learning approaches and Natural Language Processing for text classification, we suggest a simple strategy, featuring the terms according to their relative positions within the discourse sequence. For experimentations, we apply this strategy to a recent document incremental classification approach from the state-of-the-art. And, we propose an original comparison between Incremental learning and Deep learning, by comparing the incremental system with a CCN-RNN-based approach. It demonstrates that both approaches are competitive in similar contest. Vincent Poulain D'Andecy, Aurélie Joseph, Joaquín Cuenca, Jean-Marc Ogier |
ICDAR | 1 |
| 2018 | Field Extraction by Hybrid Incremental and A-Priori Structural TemplatesabstractIn this paper, we present an incremental frame-work for extracting information fields from administrative documents. First, we demonstrate some limits of the existing state-of-the-art methods such as the delay of the system efficiency. This is a concern in industrial context when we have only few samples of each document class. Based on this analysis, we propose a hybrid system combining incremental learning by means of itf-df statistics and a-priori generic models. We report in the experimental section our results obtained with a dataset of real invoices. Vincent Poulain D'Andecy, Emmanuel Hartmann, Marçal Rusiñol |
DAS | 1 |
| 2018 | InDUS: Incremental Document Understanding System Focus on Document ClassificationabstractOur objective is to propose a Document Understanding System for Digital Mailroom application which can cope with three challenges: (1) process a full workflow with high accuracy, with the constraint of a partial training; (2) minimal requirement for configuration work from expert users; (3) adapt incrementally the system in quasi real-time to continuously maximize the recall. We describe an end-to-end system based on existing incremental algorithms for both document classification and field extraction. But in this paper, we really focus on the document classification issue. The main contribution is to adapt the Incremental Growing Neural Gas (A2ING) with a dynamic incremental feature vector. Moreover, a generic Framework automatically selects textual descriptors relying on performance. The quality assessment converges the A2ING and controls the system accuracy. Vincent Poulain D'Andecy, Aurélie Joseph, Jean-Marc Ogier |
DAS | 1 |
| 2018 | Feature Selection for Document Flow SegmentationabstractIn this paper, we describe a method to restore a flow of continuous documents. The flow is a collection of consecutive scanned pages without explicit separation marks between documents. Our method is based on contextual and layout descriptors meant to specify the relationship between each pair of consecutive pages. The relationships are represented using vectors of features with boolean values indicating the presence or the absence of descriptors on concerned pages. The segmentation task therefore consists in classifying such vectors into continuities or breaks. The continuity class indicates that pages belong to the same document while the break class ends the ongoing document and starts a new one. The experimental part is based on a large collection of real administrative documents. Ahmed Hamdi, Mickaël Coustaty, Aurélie Joseph, Vincent Poulain D'Andecy, Antoine Doucet, Jean-Marc Ogier |
DAS | 4 |
| 2017 | Local Binary Patterns for Document Forgery DetectionabstractDocument forgery is an increasing problem for both the public administration and private companies. It represents substantial losses in time and economical resources. Classical solutions to this problem such as watermarks or other integrated security patterns can not be applied in general for any unknown incoming document due to the large variability on types of documents. In that scenario it is important to resort to forensic techniques to seek and analyze inconsistencies on the intrinsic features of the document image. In this paper we present a classification-based approach for forgery detection. We use uniform Local Binary Patterns (LBP) to capture discriminant texture features that are common on forged regions. Besides, we combine multiple descriptors from neighboring regions to model contextual information. Results using Support Vector Machines (SVM) for patch classification show that we are able to detect several types of forgeries in a wide range of types of documents. Francisco Cruz 0003, Nicolas Sidere, Mickaël Coustaty, Vincent Poulain D'Andecy, Jean-Marc Ogier |
ICDAR | 4 |
| 2016 | Human-Document Interaction Systems - A New Frontier for Document Image AnalysisabstractAll indications show that paper documents will not cede in favour of their digital counterparts, but will instead be used increasingly in conjunction with digital information. An open challenge is how to seamlessly link the physical with the digital -- how to continue taking advantage of the important affordances of paper, without missing out on digital functionality. This paper presents the authors' experience with developing systems for Human-Document Interaction based on augmented document interfaces and examines new challenges and opportunities arising for the document image analysis field in this area. The system presented combines state of the art camera-based document image analysis techniques with a range of complementary technologies to offer fluid Human-Document Interaction. Both fixed and nomadic setups are discussed that have gone through user testing in real-life environments, and use cases are presented that span the spectrum from business to educational applications. Dimosthenis Karatzas, Vincent Poulain D'Andecy, Marçal Rusiñol, Antoni Chica, Pere-Pau Vázquez |
DAS | 2 |
| 2016 | Entity Local Structure Graph Matching for Mislabeling CorrectionabstractThis paper proposes an entity local structure comparison approach based on inexact subgraph matching. The comparison results are used for mislabeling correction in the local structure. The latter represents a set of entity attribute labels which are physically close in a document image. It is modeled by an attributed graph describing the content and presentation features of the labels by the nodes and the geometrical features by the arcs. A local structure graph is matched with a structure model which represents a set of local structure model graphs. The structure model is initially built using a set of well chosen local structures based on a graph clustering algorithm and is then incrementally updated. The subgraph matching adopts a specific cost function that integrates the feature dissimilarities. The matched model graph is used to extract the missed labels, prune the extraneous ones and correct the erroneous label fields in the local structure. The evaluation of the structure comparison approach on 525 local structures extracted from 200 business documents achieves about 90% for recall and 95% for precision. The mislabeling correction rates in these local structures vary between 73% and 100%. Nihel Kooli, Abdel Belaïd, Aurélie Joseph, Vincent Poulain D'Andecy |
DAS | 4 |
| 2016 | A Compliant Document Image Classification System Based on One-Class ClassifierabstractDocument image classification in a professional context requires to respect some constraints such as dealing with a large variability of documents and/or number of classes. Whereas most methods deal with all classes at the same time, we answer this problem by presenting a new compliant system based on the specialization of the features and the parametrization of the classifier separately, class per class. We first compute a generalized vector of features based on global image characterization and structural primitives. Then, for each class, the feature vector is specialized by ranking the features according a stability score. Finally, a one-class K-nn classifier is trained using these specific features. Conducted experiments reveal good classification rates, proving the ability of our system to deal with a large range of documents classes. Nicolas Sidere, Jean-Yves Ramel, Sabine Barrat, Vincent Poulain D'Andecy, Saddok Kebairi |
DAS | 4 |
| 2015 | Multiresolution approach based on adaptive superpixels for administrative documents segmentation into color layersabstractAdministrative document images are usually processed in black and white what generates many problems due to the errors related to the binarization. Besides all semantic information provided by the color is lost. Document images have a rich and highly variable content. The presence of false colors and artefacts introduced by the scanning and the compression alter the segmentation of the regions. Problems arise when there is no correspondence between the point clouds which are detected in a color space and the real regions of an image. In order to help the segmentation, we propose the extraction of the main colors of an image as a set of binary layers. Due to the industrial context, our approach has to run unsupervised on a generic dataset of color administrative documents. The originality of this approach is the use of a multiresolution analysis to detect the number of colors automatically. At a low resolution, a set of local regions is obtained thanks to a SLIC-based approach which takes into account the structure of documents and which combines both colorimetric information and spatial information. Then, a merging stage is applied on each resolution separately based on the colors which have been extracted at a lower resolution. This contribution can both feed the traditional process and exploit colorimetric information. Elodie Carel, Jean-Christophe Burie, Vincent Courboulay, Jean-Marc Ogier, Vincent Poulain D'Andecy |
ICDAR | 5 |
| 2015 | One-shot field spotting on colored forms using subgraph isomorphismabstractThis paper presents an approach for spotting textual fields in commercial and administrative colored forms. We proceed by locating these fields thanks to their neighboring context which is modeled with a structural representation. First, informative zones are extracted. Second, forms are represented by graphs. In these graphs, nodes represent colored rectangular shapes while edges represent neighboring relations. Finally, the neighboring context of the queried region of interest is modeled as a graph. Subgraph isomorphism is applied in order to locate this ROI in the structural representation of a whole document. Evaluated on a 130-document image dataset, experimental results show up that our approach is efficient and that the requested information is found even if its position is changed. Maroua Hammami, Pierre Héroux, Sébastien Adam, Vincent Poulain D'Andecy |
ICDAR | 4 |
| 2013 | Field Extraction from Administrative Documents by Incremental Structural TemplatesabstractIn this paper we present an incremental framework aimed at extracting field information from administrative document images in the context of a Digital Mail-room scenario. Given a single training sample in which the user has marked which fields have to be extracted from a particular document class, a document model representing structural relationships among words is built. This model is incrementally refined as the system processes more and more documents from the same class. A reformulation of the tf-idf statistic scheme allows to adjust the importance weights of the structural relationships among words. We report in the experimental section our results obtained with a large dataset of real invoices. Marçal Rusiñol, Tayeb Benkhelfallah, Vincent Poulain D'Andecy |
ICDAR | 3 |