VLDB 2026 Research / reviewers in the wild / expert
Sebastian Wagner-Carena
dblp:206/6589
· DBLP profile ↗
3ranked-venue papers
1as first author
2since 2021 · last 2025
0000-0001-5039-1685ORCID · verified
Domains — the database's venue-derived domains; a paper can count in several
Artificial intelligence and machine learning · 2 · 1 first-author · 2 since 2021Databases, data management, data science and information retrieval · 1Graphics, computer vision, multimedia, augmented reality and games · 1
Expertise — from the expertise taxonomy: the topics of the expert's papers under the CCF categories. A weight counts papers with recency: 1 for a paper about the topic, 0.3 when the topic is its context, halved every five years.
| Artificial intelligence
1 paper |
Deep learning architectures and training · 33% Vision and language · 33% Representation and self-supervised learning · 33% | |
| Interdisciplinary, comprehensive, and emerging computing
1 paper |
Computational science and engineering · 100% |
Topics — the 5 heaviest of 6, each with the papers that count most for it
| Topic | Weight | Papers | Last | Evidence papers |
|---|---|---|---|---|
Machine learning › Deep learning architectures and training
foundation model |
0.9 | 1 | 2025 | AION-1: Omnimodal Foundation Model for Astronomical Sciences · NeurIPS 2025 |
Machine learning › Representation and self-supervised learning › representation learning › unsupervised representation learning › self-supervised representation learning
masked modeling |
0.9 | 1 | 2025 | AION-1: Omnimodal Foundation Model for Astronomical Sciences · NeurIPS 2025 |
Computer vision › Vision and language › vision-language model
multimodal large language model |
0.9 | 1 | 2025 | AION-1: Omnimodal Foundation Model for Astronomical Sciences · NeurIPS 2025 |
Computational science and engineering › astronomy
astronomical data analysis |
0.3 | 1 | 2025 | AION-1: Omnimodal Foundation Model for Astronomical Sciences · NeurIPS 2025 |
Computational science and engineering
astronomy |
0.3 | 1 | 2025 | AION-1: Omnimodal Foundation Model for Astronomical Sciences · NeurIPS 2025 |
Methods — techniques the papers use, named apart from their topics
transformer · 1.7tokenization · 1.7masked modeling · 1.7
| Year | Publication | Venue | Position |
|---|---|---|---|
| 2025 | AION-1: Omnimodal Foundation Model for Astronomical SciencesabstractWhile foundation models have shown promise across a variety of fields, astronomy lacks a unified framework for joint modeling across its highly diverse data modalities. In this paper, we present AION-1, the first large-scale multimodal foundation family of models for astronomy. AION-1 enables arbitrary transformations between heterogeneous data types using a two-stage architecture: modality-specific tokenization followed by transformer-based masked modeling of cross-modal token sequences. Trained on over 200M astronomical objects, AION-1 demonstrates strong performance across regression, classification, generation, and object retrieval tasks. Beyond astronomy, AION-1 provides a scalable blueprint for multimodal scientific foundation models that can seamlessly integrate heterogeneous combinations of real-world observations. Our model release is entirely open source, including the dataset, training script, and weights. Liam Holden Parker, François Lanusse, Jeff Shen, Ollie Liu, Tom Hehir, Leopoldo Sarra, Lucas Meyer, Micah Bowles, Sebastian Wagner-Carena, Helen Qu, Siavash Golkar, Alberto Bietti, Hatim Bourfoune, Pierre Cornette, Keiya Hirashima, Géraud Krawezik, Ruben Ohana, Nicholas Lourie, Michael McCabe, Rudy Morel, Payel Mukhopadhyay, Mariel Pettee, Kyunghyun Cho, Miles D. Cranmer, Shirley Ho |
NeurIPS | 9 |
| 2025 | A Data-Driven Prism: Multi-View Source Separation with Diffusion Model PriorsabstractIn the natural sciences, a common challenge is to disentangle distinct, unknown sources from observations. Examples of this source separation task include deblending galaxies in a crowded field, distinguishing the activity of individual neurons from overlapping signals, and separating seismic events from the ambient background. Traditional analyses often rely on simplified source models that fail to accurately reproduce the data. Recent advances have shown that diffusion models can directly learn complex prior distributions from noisy, incomplete data. In this work, we show that diffusion models can solve the source separation problem without explicit assumptions about the source. Our method relies only on multiple views, or the property that different sets of observations contain different linear transformations of the unknown sources. We show that our method succeeds even when no source is individually observed and the observations are noisy, incomplete, and vary in resolution. The learned diffusion models enable us to sample from the source priors, evaluate the probability of candidate sources, and draw from the joint posterior of our sources given an observation. We demonstrate the effectiveness of our method on a range of synthetic problems as well as real-world galaxy observations. Sebastian Wagner-Carena, Aizhan Akhmetzhanova, Sydney Erickson |
NeurIPS | 1 |
| 2018 | Simulated Annealing for JPEG QuantizationabstractJPEG is one of the most widely used image formats, but in some ways remains surprisingly unoptimized, perhaps because some natural optimizations would go outside the standard that defines JPEG. We show how to improve JPEG compression in a standard-compliant, backward-compatible manner, by finding improved default quantization tables. We describe a simulated annealing technique that has allowed us to find several quantization tables that perform better than the industry standard, in terms of both compressed size and image fidelity. Specifically, we derive tables that reduce the FSIM error by over 10% while improving compression by over 20% at quality level 95 in our tests; we also provide similar results for other quality levels. While we acknowledge our approach can in some images lead to visible artifacts under large magnification, we believe use of these quantization tables, or additional tables that could be found using our methodology, would significantly reduce JPEG file sizes with improved overall image quality. Max Hopkins, Michael Mitzenmacher, Sebastian Wagner-Carena |
DCC | 3 |