VLDB 2026 Research / reviewers in the wild / expert
Alice Rey
dblp:276/5708
· DBLP profile ↗
1ranked-venue papers
1as first author
1since 2021 · last 2025
0009-0002-9210-4124ORCID · reported
Domains — the database's venue-derived domains; a paper can count in several
Databases, data management, data science and information retrieval · 1 · 1 first-author · 1 since 2021
Expertise — from the expertise taxonomy: the topics of the expert's papers under the CCF categories. A weight counts papers with recency: 1 for a paper about the topic, 0.3 when the topic is its context, halved every five years.
| Databases, data mining, and information retrieval
1 paper |
Query processing and optimization · 100% | |
| Computer architecture, parallel and distributed computing, and storage systems
1 paper |
Storage systems · 100% |
Topics — the 3 heaviest of 3, each with the papers that count most for it
| Topic | Weight | Papers | Last | Evidence papers |
|---|---|---|---|---|
Query processing and optimization › SQL query processing
nested query processing |
0.9 | 1 | 2025 | Nested Parquet Is Flat, Why Not Use It? How To Scan Nested Data With On-the-Fly Key Generation and Joins · Proc. ACM Manag. Data 2025 |
Storage systems › data management › database storage
columnar storage |
0.9 | 1 | 2025 | Nested Parquet Is Flat, Why Not Use It? How To Scan Nested Data With On-the-Fly Key Generation and Joins · Proc. ACM Manag. Data 2025 |
Query processing and optimization
join processing |
0.3 | 1 | 2025 | Nested Parquet Is Flat, Why Not Use It? How To Scan Nested Data With On-the-Fly Key Generation and Joins · Proc. ACM Manag. Data 2025 |
Methods — techniques the papers use, named apart from their topics
on-the-fly key generation · 1.7join-based nesting reconstruction · 1.7
| Year | Publication | Venue | Position |
|---|---|---|---|
| 2025 | Nested Parquet Is Flat, Why Not Use It? How To Scan Nested Data With On-the-Fly Key Generation and JoinsabstractParquet is the most commonly used file format to store data in a columnar, binary structure. The format also supports storing nested data in this flattened columnar layout. However, many query engines either do not support nested data or process it with substantially worse performance than relational data. In this work, we close this gap and present a new way to leverage relational query engines for nested data that is stored in this flat columnar file format. Specifically, we demonstrate how to process nested Parquet files much more efficiently. Our approach does not store a copy of the data in an internal format but reads directly from the Parquet file. During query computation, the required flat columns are scanned independently and the nesting is reconstructed using joins with on-the-fly generated join keys. Our approach can be easily integrated into existing query engines to support querying nested Parquet files. Furthermore, we achieve orders of magnitude faster analytical query performance than existing solutions, which makes it a valuable addition. Alice Rey, Maximilian Rieger, Thomas Neumann 0001 |
Proc. ACM Manag. Data | 1 |