EDBT 2026 Demo / reviewers in the wild / expert
Fei Yan 0006
dblp:52/4851-6
· DBLP profile ↗
21ranked-venue papers
0as first author
19since 2021 · last 2026
0000-0001-5177-4493ORCID · conflict
Domains — the database's venue-derived domains; a paper can count in several
Artificial intelligence and machine learning · 9 · 8 since 2021Graphics, computer vision, multimedia, augmented reality and games · 8 · 8 since 2021Databases, data management, data science and information retrieval · 3 · 2 since 2021Applied, interdisciplinary, general and emerging computing · 2 · 2 since 2021Human-computer interaction and ubiquitous computing · 1 · 1 since 2021
| Year | Publication | Venue | Position |
|---|---|---|---|
| 2026 | DDL-Net: A Task-Balanced Panoptic Perception Network with Channel-Reorganized Attention in Autonomous DrivingabstractMulti-task perception in traffic scenes requires unified representations for object recognition, region understanding, and structural reasoning. However, heterogeneous tasks such as detection and segmentation impose conflicting optimization objectives on shared features, leading to representation imbalance and reduced robustness. We propose DDL-Net, a task-balanced panoptic perception network that addresses this issue through coordinated representation learning and task-aware decoding, introducing a Channel-Rearranging Attention Interaction (CRAI) module built upon Channel Rearranging Boosting Attention (CRBA) units to reduce attention redundancy and enhance informative channel responses, improving representation of small and degraded targets. In addition, a task-specific decoding mechanism combining the Task-Guided Feature Balancing Module (TGBM) and Frequency-Aware Task-Interaction Guided Decoder (FTIG-Decoder) aligns shared features with task semantics, enabling balanced performance across detection and segmentation tasks. Experiments on BDD100K validate the effectiveness and robustness of the proposed design. Bowei Fang, Yunyi Tang, Wentao Mu, Wenbo Liu 0006, Fei Yan 0006, Tao Deng 0002 |
ICMR | 5 |
| 2026 | Secure Consensus Control of Nonlinear Multiagent Systems Based on Memory Adaptive ProtocolabstractThis article investigates memory-based adaptive event-triggered (MBAET) consensus control for nonlinear multiagent systems (MASs) subject to multiple cyberattacks. Amultiattack model is formulated, incorporating deception and irregular denial-of-service (DoS) attacks, with a focus on the frequency and duration of DoS attacks. To reduce communication overhead, an improved MBAET mechanism is proposed in which a buffer at the triggering node stores historical information, allowing previously transmitted data to be exploited in the trigger design. Moreover, the designed threshold adaptation law adjusts the triggering threshold online based on stored historical data and prescribed performance requirements, to regulate the convergence process in real time. ALyapunov function incorporating communication delays and sampling instants is constructed. During DoS attack intervals, only locally available system information is utilized, which relaxes the associated functional constraints. By further adopting a piecewise Lyapunov functional, we derive stochastic mean-square consensus conditions for MASs under multiple attacks, covering both attack and nonattack intervals. The MBAET matrix parameters and controller gains are obtained using linear matrix inequality (LMI) techniques. Finally, the effectiveness and advantages of the proposed method in enhancing attack resilience and improving resource efficiency are demonstrated through a practical simulation based on an F-404 aircraft engine system. Kangyue Chen, Fei Yan 0006 |
IEEE Trans. Syst. Man Cybern. Syst. | 2 |
| 2025 | SalM²: An Extremely Lightweight Saliency Mamba Model for Real-Time Cognitive Awareness of Driver AttentionabstractDriver attention recognition in driving scenarios is a popular direction in traffic scene perception technology. It aims to understand human driver attention to focus on specific targets/objects in the driving scene. However, traffic scenes contain not only a large amount of visual information but also semantic information related to driving tasks. Existing methods lack attention to the actual semantic information present in driving scenes. Additionally, the traffic scene is a complex and dynamic process that requires constant attention to objects related to the current driving task. Existing models, influenced by their foundational frameworks, tend to have large parameter counts and complex structures. Therefore, this paper proposes a real-time saliency Mamba network based on the latest Mamba framework. As shown in Figure 1, our model uses very few parameters (0.08M, only 0.09~11.16% of other models), while maintaining SOTA performance or achieving over 98% of the SOTA model's performance. Wentao Mu, Wenbo Liu 0006, Fei Yan 0006, Tao Deng 0002 |
AAAI | 5 |
| 2025 | VP-YOLO: Robust Vehicle-Pedestrian Detection in Challenging Traffic Scenarios via A Human Visual Perception-Inspired NetworkabstractIntelligent vehicles need to provide rational driving strategies for assisted driving systems based on driving scenarios. Since pedestrians and vehicles are the main players in these scenarios, accurate detection and localization of pedestrians and vehicles are crucial for intelligent driving systems to make reliable decisions in dynamic environments. However, existing pedestrian and vehicle detection models often lack robustness under dynamic and complex traffic conditions, resulting in missed detections and false alarms, which pose significant safety risks. To address this problem, we categorize complex traffic scenarios into three typical challenges: long-distance, truncation, and occlusion, and focus on designing a novel enhancement stage to make the model more robust to these challenges. In this enhancement stage, inspired by human visual perception, we design a Visual Attention Module (VAM). This module can gather high-quality horizontal and vertical spatial features and efficiently interact between horizontal and vertical spatial features, enhancing the model’s perceptual ability by mimicking optic chiasm. Additionally, we use a Feature Reconstruction Module (FRM) to reduce redundant information in the feature maps and enhance the model’s inference ability. We conduct comprehensive experiments on the KITTI benchmark and Cityscapes dataset, and the experimental results demonstrate that our algorithm achieves state-of-the-art performance in various challenging scenarios. Wenbo Liu 0006, Tao Deng 0002, Fei Yan 0006 |
ICASSP | 3 |
| 2025 | HID-NAS: A Novel Neural Architecture Search Pipeline for High Information Density DataabstractNeural Architecture Search (NAS) is a core component of automated machine learning, enabling the automatic discovery of task-specific architectures. However, directly employing raw data for NAS faces challenges in terms of large data volume, low search efficiency, and potential data privacy concerns. To address these limitations, we condense large-scale target datasets into high information-density proxy datasets. And we propose a novel pipeline, HID-NAS, which skillfully exploits High Information Density(HID) data for efficient neural architecture search. At the algorithmic level, our algorithm focuses on the inherent problem of DARTS-derived algorithms, i.e., the discrepancy between trained supernet and selected architectures. We pioneer the phenomenon of "Selection Error" that occurs during architecture selection and introduce an attention mechanism that focuses on candidate edges to improve the accuracy of pipeline-selected architectures. Furthermore, we also formulate a regularization mechanism that employs the high information density data of the proxy dataset to transform the architectural parameter updates into an adversarial game and adjust the strengths through a normalization mechanism. These three key components – attention, regularization, and normalization – allow our pipeline to efficiently identify high-quality models using the distilled dataset. Experimental results demonstrate a significant reduction in search time, achieving high-quality models within just two minutes. Wenbo Liu 0006, Tao Deng 0002, Fei Yan 0006 |
ICASSP | 3 |
| 2025 | Ultra-Lightweight Thyroid Puncture Positioning Detection Guided by Nodule Location
Shengqi Chen 0004, Yi Huang 0022, Chengfan Yang, Buyun Ma, Fei Yan 0006, Yang Chen 0060, Tao Deng 0002 |
PRCV (13) | 5 |
| 2025 | Driving Fixation Prediction for Clear-to-Adverse Weather Scenes via Adversarial Unsupervised Domain Adaptation
Wentao Mu, Wenbo Liu 0006, Yanghua Zhang, Fei Yan 0006, Tao Deng 0002 |
PRCV (12) | 6 |
| 2025 | Improving vehicle detection accuracy in complex traffic scenes through context attention and multi-scale feature fusion module
Wenbo Liu 0006, Binglin Zhao, Tao Deng 0002, Fei Yan 0006 |
Appl. Intell. | 5 |
| 2025 | Continuous-Discrete Alignment Optimization for efficient differentiable neural architecture searchabstractDifferential Architecture Search (DARTS) has become a prominent technique for neural architecture search in recent years. Despite its merits, the issue of discretization discrepancy within DARTS still necessitates further exploration, as it can degrade in performance. In this paper, we introduce a novel algorithm termed Continuous–Discrete Alignment Optimization (DARTS-CDAO), designed to address the discretization discrepancy and thereby enhance the robustness and generalization capabilities of the discovered neural architectures. Our proposed DARTS-CDAO algorithm seamlessly integrates the discretization process into the training phase of the architecture parameters, thereby bolstering the search algorithm’s adaptability to the inherent discretization processes. Specifically, our methodology commences by formalizing the process of architecture parameter discretization. Subsequently, we introduce a coarse gradient weighting algorithm that is employed to update the architecture parameters, effectively minimizing the divergence between the representation of continuous and discrete parameters. Rigorous theoretical analysis, coupled with extensive experimental outcomes, substantiates that our proposed approach can elevate the performance of the searched models. Notably, this enhancement is achieved without incurring additional search time, rendering DARTS more robust and endowed with a heightened capacity for generalization. Wenbo Liu 0006, Jia Wu 0005, Tao Deng 0002, Fei Yan 0006 |
Eng. Appl. Artif. Intell. | 4 |
| 2025 | VP-YOLO: A human visual perception-inspired robust vehicle-pedestrian detection model for complex traffic scenarios
Wenbo Liu 0006, Xiaoyun Qiao, Tao Deng 0002, Fei Yan 0006 |
Expert Syst. Appl. | 5 |
| 2025 | A new pipeline with ultimate search efficiency for neural architecture search
Wenbo Liu 0006, Xiaoyun Qiao, Tao Deng 0002, Fei Yan 0006 |
Neural Networks | 5 |
| 2025 | VP2Net: Visual Perception-Inspired Network for Exploring the Causes of Drivers' Attention ShiftabstractWith the rapid development of autonomous driving technology, the recognition/understanding of driving events has become increasingly important for improving road safety. Existing methods for recognizing driving events rely solely on the inherent features of driving scenes, lacking real-time modeling of driver attention and the integration of driver attention for understanding driving events. Research has shown that understanding driver attention will be beneficial for subsequent analysis of driving events. We propose the attention-based driving event dataset (ADED), which includes rich driving scenes, eye movement data, reasons for attention shifts, and event time windows. It enables the use of prior information about driver attention to guide the recognition of driving events. Based on our dataset, we propose a visual dual-perception network, named VP2Net, to explore the reasons behind driver attention shifts. The goal of VP2Net is to use driver attention to guide the recognition of driving events. Inspired by the human visual dual cognition process mechanism, we build a bottom-up sequential information encoding branch for extracting spatio-temporal low-level information in the driving scene. Additionally, we establish a top-down attention perceptual encoding branch that simulates the driver’s high-level visual cognitive process. It not only captures the driver’s spatial attention allocation (“where to focus”) but also performs a temporal dimensional perceptual enhancement (“when to focus”), allowing us to extract the driver’s spatial attention enhancement information. We use the driver’s spatial attention enhancement information to guide the fusion of spatio-temporal information of the driving scene and selectively highlight the core objects/areas in the current driving task/event. Finally, we compare our proposed model with other SOTA networks and visualize the results of the key components of the model. Our code is available athttps://github.com/zhao-chunyu/VP2Net Tao Deng 0002, Pengcheng Du, Wenbo Liu 0006, Yi Huang 0022, Fei Yan 0006 |
IEEE Trans. Intell. Transp. Syst. | 6 |
| 2024 | DARTS-CGW: Research on Differentiable Neural Architecture Search Algorithm Based on Coarse Gradient Weighting
Wenbo Liu 0006, Tao Deng 0002, Rui An, Fei Yan 0006 |
PRCV (3) | 4 |
| 2023 | What Causes a Driver's Attention Shift? A Driver's Attention-Guided Driving Event Recognition ModelabstractDespite much effort to try to research driver's spatial attention allocation in driving situations, the computer vision community rarely focuses on what causes driver's attention shifts. In this paper, we built an attention-based driving event dataset (ADED) constructed from the attention distributed on the traffic participants or elements and proposed a model using driver's attention as the guidance to better recognize the events that lead driver's attention shifts. We relabeled and redivided BDD-A, a driver attention dataset in critical traffic situations, into six different semantic categories of driving events. The new dataset is introduced for driving event recognition. In addition, we proposed a special driving event recognition model (called DER-Net) with driver's attention guidance to recognize the event causes a driver's attention shift. In DER-Net, a driver's attention-guided (DAG) branch is constructed to consider the driver's spatiotemporal attention information. The proposed model achieves a superior performance compared to other state-of-the-art models in action recognition. In ablation study, many experiments are conducted to discuss proper length of the sequence to input the model, the optimal criterion to find out the frame when driver's attention shifts and the appropriate function to quantify different attention maps in neural network. Pengcheng Du, Tao Deng 0002, Fei Yan 0006 |
IJCNN | 3 |
| 2022 | TARConvGRU: A Cross-dimension Spatiotemporal Model for Lane DetectionabstractSpatiotemporal information plays a critical role in the autonomous driving environment. In the task of lane detection, the features extracted by the existing spatiotemporal work from a video or consecutive frames contain a lot of irrelevant information, which reduces the performance of the model to a certain extent. To address the trouble, we proposed a novel lane detection model with the TARConvGRU module. This module can help our model focus on the feature representation and feature location of lane lines by channel and spatial operations and consciously guides some computing resources tend to the most likely lane line features by capturing the cross-dimension information interaction. Furthermore, we demonstrate the effectiveness of the proposed model on three popular lane detection benchmarks: TuSimple, Unsupervised Labeled Lane Markers (Unsupervised LLAMAS), and Dynamic Vision Sensor Dataset (DET). The experiments show that our model achieves competitive results compared with other state-of-the-art models. In addition, we discuss some significant problems encountered in the designing process in detail in the ablation study. Tao Deng 0002, Fei Yan 0006 |
IJCNN | 3 |
| 2022 | Fastest containment control of discrete-time multi-agent systems using static linear feedback protocol
Fei Yan 0006, Tao Feng 0006, Tao Deng 0002, Yue Zhao 0004 |
Inf. Sci. | 2 |
| 2022 | Lane Detection Model Based on Spatio-Temporal Network With Double Convolutional Gated Recurrent UnitsabstractLane detection is one of the indispensable and key elements of self-driving environmental perception. Many lane detection models have been proposed, solving lane detection under challenging conditions, including intersection merging and splitting, curves, boundaries, occlusions and combinations of scene types. Nevertheless, lane detection will remain an open problem for some time to come. The ability to cope well with those challenging scenes impacts greatly the applications of lane detection on advanced driver assistance systems (ADASs). In this paper, a spatio-temporal network with double Convolutional Gated Recurrent Units (ConvGRUs) is proposed to address lane detection in challenging scenes. Both of ConvGRUs have the same structures, but different locations and functions in our network. One is used to extract the information of the most likely low-level features of lane markings. The extracted features are input into the next layer of the end-to-end network after concatenating them with the outputs of some blocks. The other one takes some continuous frames as its input to process the spatio-temporal driving information. Extensive experiments on the large-scale TuSimple lane marking challenge dataset and Unsupervised LLAMAS dataset demonstrate that the proposed model can effectively detect lanes in the challenging driving scenes. Our model can outperform the state-of-the-art lane detection models. Tao Deng 0002, Fei Yan 0006, Wenbo Liu 0006 |
IEEE Trans. Intell. Transp. Syst. | 3 |
| 2021 | Driving Video Fixation Prediction Model Via Spatio-Temporal Networks and Attention GatesabstractDriving fixation prediction is becoming an essential research problem in human-like driving systems or advanced driver assistance systems (ADAS) in a dynamic driving environment. However, it is still a lack of driving video fixation prediction models that can dynamically predict drivers’ fixational locations. In this work, we propose a driving video fixation prediction model via spatio-temporal networks and attention gates method, named as DSTANet, to predict drivers’ attention in the dynamic driving videos. The spatial and temporal driving information are both considered in DSTANet by convolutional long short-term memory (ConvLSTM). In addition, the human attention mechanism is designed to filter some driving-irrelevant information via the attention gates (AGs). The experimental results indicate that the proposed DSTANet outperforms the state-of-the-art saliency models and predicts drivers’ attentional spatial locations more accurately. Furthermore, the time-series prediction results show that DSTANet is more robust and coherent than others and includes more temporal information. Tao Deng 0002, Fei Yan 0006 |
ICME | 2 |
| 2021 | Optimal distributed cooperative control for multi-agent systems with constrains on convergence speed and control input
Tao Feng 0006, Fei Yan 0006 |
Neurocomputing | 4 |
| 2019 | Adaptive sliding mode fault-tolerant control for type-2 fuzzy systems with distributed delays
Yue Zhao 0004, Fei Yan 0006, Yi Shen 0001 |
Inf. Sci. | 3 |
| 2018 | Hierarchical Temporal Memory method for time-series-based anomaly detection
Jia Wu 0005, Weiru Zeng, Fei Yan 0006 |
Neurocomputing | 3 |