Xiaohan Fei

dblp:172/1373 · DBLP profile ↗
← Back
4ranked-venue papers
3as first author
1since 2021 · last 2021
0000-0002-1030-2286ORCID · corroborated

Domains — the database's venue-derived domains; a paper can count in several

Artificial intelligence and machine learning · 4 · 3 first-author · 1 since 2021Graphics, computer vision, multimedia, augmented reality and games · 4 · 3 first-author · 1 since 2021

Expertise — from the expertise taxonomy: the topics of the expert's papers under the CCF categories. A weight counts papers with recency: 1 for a paper about the topic, 0.3 when the topic is its context, halved every five years.

Artificial intelligence
4 papers
3D vision · 67% Robot navigation and mapping · 33%
Theoretical computer science
1 paper
Algorithms and data structures · 100%

Topics — the 9 heaviest of 9, each with the papers that count most for it

TopicWeightPapersLastEvidence papers
Computer vision › 3D vision
camera calibration
0.512021
Single View Physical Distance Estimation using Human Pose · ICCV 2021
Computer vision › 3D vision
depth estimation
0.512021
Single View Physical Distance Estimation using Human Pose · ICCV 2021
Robotics › Robot navigation and mapping
SLAM
0.322018
A Simple Hierarchical Pooling Data Structure for Loop Closure · ECCV (3) 2016
Visual-Inertial Object Detection and Mapping · ECCV (11) 2018
Robotics › Robot navigation and mapping › visual odometry
visual-inertial odometry
0.312018
Visual-Inertial Object Detection and Mapping · ECCV (11) 2018
Computer vision › 3D vision
3d object detection
0.312017
Visual-Inertial-Semantic Scene Representation for 3D Object Detection · CVPR 2017
Computer vision › 3D vision › 3d scene modeling
scene representation
0.312017
Visual-Inertial-Semantic Scene Representation for 3D Object Detection · CVPR 2017
Computer vision › 3D vision › 3d scene modeling › scene representation
semantic scene representation
0.312017
Visual-Inertial-Semantic Scene Representation for 3D Object Detection · CVPR 2017
Robotics › Robot navigation and mapping › SLAM
loop closure detection
0.212016
A Simple Hierarchical Pooling Data Structure for Loop Closure · ECCV (3) 2016
Algorithms and data structures › data structure design
hierarchical data structures
0.112016
A Simple Hierarchical Pooling Data Structure for Loop Closure · ECCV (3) 2016

Methods — techniques the papers use, named apart from their topics

loop closure · 0.5human pose priors · 0.5hierarchical pooling · 0.5direct formulation · 0.5visual-inertial fusion · 0.3object detection · 0.3localization and mapping filter · 0.3inertial sensing · 0.3convolutional neural network · 0.3
YearPublicationVenuePosition
2021 Single View Physical Distance Estimation using Human Pose
abstract
We propose a fully automated system that simultaneously estimates the camera intrinsics, the ground plane, and physical distances between people from a single RGB image or video captured by a camera viewing a 3-D scene from a fixed vantage point. To automate camera calibration and distance estimation, we leverage priors about human pose and develop a novel direct formulation for pose-based auto-calibration and distance estimation, which shows state-of-the-art performance on publicly available datasets. The proposed approach enables existing camera systems to measure physical distances without needing a dedicated calibration process or range sensors, and is applicable to a broad range of use cases such as social distancing and workplace safety. Furthermore, to enable evaluation and drive research in this area, we contribute to the publicly available MEVA dataset with additional distance annotations, resulting in "MEVADA" – an evaluation benchmark for the pose-based auto-calibration and distance estimation problem.
Xiaohan Fei, Henry Wang, Lin Lee Cheong, Joseph Tighe
ICCV1
2018 Visual-Inertial Object Detection and Mapping
Xiaohan Fei, Stefano Soatto
ECCV (11)1
2017 Visual-Inertial-Semantic Scene Representation for 3D Object Detection
abstract
We describe a system to detect objects in three-dimensional space using video and inertial sensors (accelerometer and gyrometer), ubiquitous in modern mobile platforms from phones to drones. Inertials afford the ability to impose class-specific scale priors for objects, and provide a global orientation reference. A minimal sufficient representation, the posterior of semantic (identity) and syntactic (pose) attributes of objects in space, can be decomposed into a geometric term, which can be maintained by a localization-and-mapping filter, and a likelihood function, which can be approximated by a discriminatively-trained convolutional neural network The resulting system can process the video stream causally in real time, and provides a representation of objects in the scene that is persistent: Confidence in the presence of objects grows with evidence, and objects previously seen are kept in memory even when temporarily occluded, with their return into view automatically predicted to prime re-detection.
Jingming Dong, Xiaohan Fei, Stefano Soatto
CVPR2
2016 A Simple Hierarchical Pooling Data Structure for Loop Closure
Xiaohan Fei, Konstantine Tsotsos, Stefano Soatto
ECCV (3)1