VLDB 2026 Research / reviewers in the wild / expert
Shuai Han 0005
dblp:59/9242-5
· DBLP profile ↗
13ranked-venue papers
3as first author
12since 2021 · last 2026
—ORCID · conflict
Domains — the database's venue-derived domains; a paper can count in several
Artificial intelligence and machine learning · 7 · 2 first-author · 6 since 2021Databases, data management, data science and information retrieval · 6 · 1 first-author · 6 since 2021
| Year | Publication | Venue | Position |
|---|---|---|---|
| 2026 | TATRC: Triple Actor-Critic Structure with Regularization for better performance
Taihong Zhong, Shuai Han 0005, Zehong Long, Shuai Lü 0001, Junhong Wu |
Inf. Process. Manag. | 2 |
| 2026 | Influence of Gaussian distribution on performance metrics in continuous reinforcement learning
Ruikai Zhou, Taihong Zhong, Wenbo Zhu 0003, Shuai Han 0005, Shuai Lü 0001 |
Inf. Process. Manag. | 4 |
| 2025 | VCSAP: Online reinforcement learning exploration method based on visitation count of state-action pairs
Ruikai Zhou, Wenbo Zhu 0003, Shuai Han 0005, Meng Kang, Shuai Lü 0001 |
Neural Networks | 3 |
| 2024 | Mixed experience sampling for off-policy reinforcement learning
Jiayu Yu, Jingyao Li 0003, Shuai Lü 0001, Shuai Han 0005 |
Expert Syst. Appl. | 4 |
| 2024 | Explorer-Actor-Critic: Better actors for deep reinforcement learning
Shuai Han 0005, Shuai Lü 0001 |
Inf. Sci. | 2 |
| 2023 | Guided deterministic policy optimization with gradient-free policy parameters information
Shuai Han 0005, Xiaoyu Gong, Shuai Lü 0001 |
Expert Syst. Appl. | 3 |
| 2023 | Entropy regularization methods for parameter space exploration
Shuai Han 0005, Wenbo Zhou 0003, Shuai Lü 0001, Xiaoyu Gong |
Inf. Sci. | 1 |
| 2022 | NROWAN-DQN: A stable noisy network with noise reduction and online weight adjustment for exploration
Shuai Han 0005, Wenbo Zhou 0003, Shuai Lü 0001 |
Expert Syst. Appl. | 1 |
| 2022 | Sampling diversity driven exploration with state difference guidance
Shuai Han 0005, Shuai Lü 0001, Meng Kang |
Expert Syst. Appl. | 2 |
| 2022 | Proximal policy optimization via enhanced exploration efficiency
Shuai Han 0005, Shuai Lü 0001 |
Inf. Sci. | 3 |
| 2021 | Recruitment-imitation mechanism for evolutionary reinforcement learning
Shuai Lü 0001, Shuai Han 0005, Wenbo Zhou 0003 |
Inf. Sci. | 2 |
| 2021 | Regularly updated deterministic policy gradient algorithm
Shuai Han 0005, Wenbo Zhou 0003, Shuai Lü 0001, Jiayu Yu |
Knowl. Based Syst. | 1 |
| 2020 | Deep Recurrent Deterministic Policy Gradient for Physical Control
Shuai Han 0005, Zhiruo Zhang, Lefan Li, Shuai Lü 0001 |
ICANN (2) | 2 |