Simon Doll

dblp:331/6698 · DBLP profile ↗
← Back
2ranked-venue papers
2as first author
2since 2021 · last 2024
0000-0001-7985-0830ORCID · corroborated

Domains — the database's venue-derived domains; a paper can count in several

Artificial intelligence and machine learning · 2 · 2 first-author · 2 since 2021Graphics, computer vision, multimedia, augmented reality and games · 2 · 2 first-author · 2 since 2021

Expertise — from the expertise taxonomy: the topics of the expert's papers under the CCF categories. A weight counts papers with recency: 1 for a paper about the topic, 0.3 when the topic is its context, halved every five years.

Artificial intelligence
2 papers
3D vision · 52% Autonomous driving · 32% Motion planning and robot control · 16%

Topics — the 7 heaviest of 7, each with the papers that count most for it

TopicWeightPapersLastEvidence papers
Computer vision › 3D vision › 3d scene modeling › scene representation
dynamic scene representation
0.812024
Dualad: Disentangling the Dynamic and Static World for End-to-End Driving · CVPR 2024
Robotics › Autonomous driving
end-to-end driving
0.812024
Dualad: Disentangling the Dynamic and Static World for End-to-End Driving · CVPR 2024
Robotics › Motion planning and robot control › robot control
motion compensation
0.812024
Dualad: Disentangling the Dynamic and Static World for End-to-End Driving · CVPR 2024
Robotics › Autonomous driving
perception
0.812024
Dualad: Disentangling the Dynamic and Static World for End-to-End Driving · CVPR 2024
Computer vision › 3D vision
3d object detection
0.612022
SpatialDETR: Robust Scalable Transformer-Based 3D Object Detection From Multi-view Camera Images With Global Cross-Sensor Attention · ECCV (39) 2022
Computer vision › 3D vision › 3d object detection › image-based 3d object detection
multi-view 3d object detection
0.612022
SpatialDETR: Robust Scalable Transformer-Based 3D Object Detection From Multi-view Camera Images With Global Cross-Sensor Attention · ECCV (39) 2022
Computer vision › 3D vision › 3d object detection
transformer-based 3d object detection
0.612022
SpatialDETR: Robust Scalable Transformer-Based 3D Object Detection From Multi-view Camera Images With Global Cross-Sensor Attention · ECCV (39) 2022

Methods — techniques the papers use, named apart from their topics

latent representation learning · 0.8cross-attention · 0.8transformer · 0.6cross-sensor attention · 0.6
YearPublicationVenuePosition
2024 Dualad: Disentangling the Dynamic and Static World for End-to-End Driving
abstract
State-of-the-art approaches for autonomous driving integrate multiple sub-tasks of the overall driving task into a single pipeline that can be trained in an end-to-end fashion by passing latent representations between the different modules. In contrast to previous approaches that rely on a unified grid to represent the belief state of the scene, we propose dedicated representations to disentangle dynamic agents and static scene elements. This allows us to explicitly compensate for the effect of both ego and object motion between consecutive time steps and to flexibly propagate the belief state through time. Furthermore, dynamic objects can not only attend to the input camera images, but also directly benefit from the inferred static scene structure via a novel dynamic-static cross-attention. Extensive experiments on the challenging nuScenes benchmark demonstrate the benefits of the proposed dual-stream design, especially for modelling highly dynamic agents in the scene, and highlight the improved temporal consistency of our approach. Our method titled DualAD not only outperforms independently trained single-task networks, but also improves over previous state-of-the-art end-to-end models by a large margin on all tasks along the functional chain of driving.
Simon Doll, Niklas Hanselmann, Lukas Schneider, Richard Schulz, Marius Cordts, Markus Enzweiler, Hendrik P. A. Lensch
CVPR1
2022 SpatialDETR: Robust Scalable Transformer-Based 3D Object Detection From Multi-view Camera Images With Global Cross-Sensor Attention
Simon Doll, Richard Schulz, Lukas Schneider, Viviane Benzin, Markus Enzweiler, Hendrik P. A. Lensch
ECCV (39)1