Demonstration venue · read-only. Every page can be browsed; the buttons that would change it are switched off. Create an account to run TaxoReview on your own data.
EDBT 2026 Demo / recruiting

Reviewers in the wild

The CS expert database: 1,777,551 experts · 2,837,695 papers in CCF A–C and ICORE A*–C venues · 10 venue-derived domains · snapshot 2026-09-22.

Recruitment bar

Who qualifies as a reviewer for this venue: pick venues from the rank lists, then set the junior and senior requirements. Each role keeps its own saved bar, which diagnosis and pool planning also use.
Saved reviewer bar: SIGMOD, VLDB, ICDE, PODS, SIGKDD · junior 2+ first-author in 5y or senior 4+ papers in 5y
1Venues the expert should have published in
The top venues whose papers the expertise taxonomy profiles: ranked A by CCF, A* by ICORE or A by TH-CPL. Expertise profiles are built from papers in these venues only.
No venue ticked: any listed venue counts
Computer architecture / parallel and distributed computing / storage systems 19
CCF A
CCF B
Computer networks 14
CCF A
CCF B
Network and information security 10
CCF A
CCF B
Software engineering / systems software / programming languages 15
CCF A
CCF C
Databases / data mining / content retrieval 12
CCF A
CCF B
Theory of computer science 12
CCF A
CCF B
Computer graphics and multimedia 9
CCF A
CCF B
Artificial intelligence 19
CCF A
CCF B
Human-computer interaction and ubiquitous computing 5
CCF A
CCF B
Interdisciplinary / comprehensive / emerging 8
CCF A
CCF B
4601Applied computing 1
ICORE B
4602Artificial intelligence 10
ICORE A*
4603Computer vision and multimedia computation 7
ICORE A*
4604Cybersecurity and privacy 7
ICORE A*
ICORE A
4605Data management and data science 10
ICORE A*
ICORE A
4606Distributed computing and systems software 21
ICORE A*
ICORE A
ICORE B
4607Graphics, augmented reality and games 3
ICORE A*
4608Human-centred computing 8
ICORE A*
ICORE A
4611Machine learning 9
ICORE A*
ICORE A
4612Software engineering 12
ICORE A*
ICORE A
ICORE B
ICORE C
4613Theory of computation 11
ICORE A*
ICORE A
ICORE B
CSEComputer Systems Engineering 8
ICORE A*
ICORE A
Listed are the profiled venues with papers since 2020. A venue under a lower rank here is top-tier in the list named on it. A venue ranked by both catalogues, or filed under several fields, is one entry: ticking any copy ticks them all. Rank labels are the catalogues', not a quality judgement.
2Publication requirement
at least papers as first author in the recent years
at least papers as any author in the recent years
“Recent k years” means papers dated 2026−k+1 or later; 2026 is incomplete in the snapshot and a few journals already date papers 2027. First author = first-listed, not corresponding.

Experts by research domain — the database's own venue-derived domains; a paper can belong to several; active = at least one paper since 2021

Experts by expertise

The expertise taxonomy: 90,661 topics extracted from 634,694 papers of 537,692 experts, organised under the ten CCF categories and 168 areas. Pick a category, then narrow down to a topic; every level lists its experts by weight.
Artificial intelligence›Machine learning›Reinforcement learning area 16,870 papers · 30,121 experts
Sub-topics 244 · with their number of papers; shaded topics divide further
multi-agent reinforcement learning 1,511exploration 952imitation learning 794multi-armed bandit 732offline reinforcement learning 732model-based reinforcement learning 708deep reinforcement learning 706markov decision process 613policy optimization 592regret minimization 497bandit 431hierarchical reinforcement learning 425reinforcement learning from human feedback 371policy learning 359actor-critic methods 324safe reinforcement learning 282value-based reinforcement learning 271temporal difference learning 242off-policy evaluation 224sample efficiency 207off-policy reinforcement learning 197value function approximation 194meta-reinforcement learning 176policy evaluation 149value function estimation 148reward learning 139continuous control 137thompson sampling 137reward design 136preference learning 128constrained reinforcement learning 126goal-conditioned reinforcement learning 121policy search 121partially observable reinforcement learning 115multi-objective reinforcement learning 110robust reinforcement learning 110function approximation 107multi-task reinforcement learning 104model-free reinforcement learning 101transfer learning in reinforcement learning 86generalization in reinforcement learning 57bayesian reinforcement learning 56dynamic programming 55reinforcement learning theory 54reinforcement learning for control 42maximum entropy reinforcement learning 40action selection 34reinforcement learning for combinatorial optimization 32policy adaptation 31unsupervised reinforcement learning 30delayed feedback 29sparse reward reinforcement learning 28agent evaluation 27curriculum reinforcement learning 25episodic reinforcement learning 25non-stationary reinforcement learning 25sparse reward 25reward maximization 24partial observability 21human-in-the-loop reinforcement learning 20reinforcement learning for reasoning 19relational reinforcement learning 19human feedback 18value function 18online control 17reinforcement learning environment 17representation learning for control 17differentiable simulation 15on-policy reinforcement learning 15stochastic control 15bellman equation 14continuous-time reinforcement learning 14stochastic environment 14long-horizon tasks 12policy diversity 12reinforcement learning with function approximation 12trajectory modeling 12benchmark design 11cost function learning 11human behavior modeling 11multi-turn reinforcement learning 11reinforcement learning for recommendation 11self-improving agent 11causal reinforcement learning 10hybrid reinforcement learning 10kernel-based reinforcement learning 10large action space 10large-scale reinforcement learning 10LLM agent training 10safety constraints 10self-supervised reinforcement learning 10embodied control 7ensemble reinforcement learning 7knowledge-based reinforcement learning 7memory architectures 7online decision making 6action space design 5policy composition 5reinforcement learning for NLP 5regularization for reinforcement learning 4population-based learning 3
Show 143 smaller topics (under 10 papers each)
embodied agent training 9learning automata 9non-stochastic control 9reinforcement learning for healthcare 9active inference 8data augmentation for reinforcement learning 8delayed reinforcement learning 8generalist agents 8offline learning 8reinforcement learning for vision 8auto-bidding 7discount factor 7efficient reinforcement learning 7hindsight relabeling 7learning from failure 7preference feedback 7real-world reinforcement learning 7reinforcement learning from process rewards 7two-timescale stochastic approximation 7adaptive discretization 6belief state 6competitive analysis 6goal-reaching tasks 6language-conditioned policy 6oracle-efficient algorithms 6policy selection 6reinforcement learning training 6reward uncertainty 6sparse-reward environments 6strategic decision-making 6target network 6task inference 6action elimination 5automated reinforcement learning 5autonomous reinforcement learning 5behavioral prior 5cognitive control 5continuous state-action spaces 5dispatching strategy 5factored reinforcement learning 5heavy-tailed reward 5oculomotor control 5outcome-based reinforcement learning 5parallel reinforcement learning 5policy training 5reinforcement learning for structured prediction 5reinforcement learning library 5rich observations 5self-evolution 5sequential experimental design 5simulation-based training 5trajectory alignment 5adaptive data collection 4adaptive optimal control 4adaptive policy 4agent behavior analysis 4bilevel reinforcement learning 4board game playing 4dynamic weighting 4empirical likelihood 4episodic learning 4experience reuse 4experience-based learning 4fleet management 4goal-conditioned control 4goal-driven learning 4guided reinforcement learning 4hybrid-action reinforcement learning 4hyperparameter scheduling 4interactive task learning 4learned objective functions 4low-rank structure 4planning and learning 4plasticity preservation 4policy estimation 4policy generation 4policy initialization 4real-time learning 4reinforcement learning for optimization 4reward function 4risk-sensitive policy 4safety-critical decision making 4semiparametric efficient estimation 4sequential task learning 4spatial crowdsourcing task assignment 4specification inference 4switching costs 4task composition 4textual reinforcement learning 4two-player game 4undiscounted reinforcement learning 4variance minimization 4affordances 3agent-environment interaction 3algorithmic alignment 3behavior foundation models 3behavior learning 3biological sequence design 3biologically plausible reinforcement learning 3context detection 3continuous-time control 3delayed rewards 3demonstration dataset 3diversity optimization 3event-based reinforcement learning 3goal representation 3high-dimensional dynamical systems 3image denoising 3iterative self-learning 3koopman operator learning 3large-scale decision making 3learning from user feedback 3learning progress prediction 3long-horizon control 3market simulation 3markov reward process 3modular policy 3multi-step lookahead 3multimodal reinforcement learning 3neuromodulation 3observational learning 3personalized reinforcement learning 3player evaluation 3policy alignment 3policy computation 3preference specification 3rare-event simulation 3reinforcement learning for generative models 3reinforcement learning for vision-language tasks 3reinforcement learning from reward models 3rollout budget allocation 3runtime adaptation 3sample allocation 3sparse reward learning 3state augmentation 3stochastic dominance 3strategy selection 3structure detection 3supervised learning for RL 3tabular reinforcement learning 3trajectory augmentation 3value prediction 3world simulator 3
Experts on Reinforcement learning 30,121 · by weight: papers on the topic counted with recency (1 for a paper about it, 0.3 as context, halved every five years), summed over the area's topics
ExpertWeight Topics held All papersLastORCID
Sergey Levine
dblp:80/7594
228.2 154 371 · 196 since 2021 2026 verified
Zhuoran Yang
dblp:172/1424
125.1 90 134 · 94 since 2021 2026 conflict
Pieter Abbeel
dblp:a/PieterAbbeel
122.9 116 330 · 127 since 2021 2026 verified
Yang Yu 0001
dblp:46/2181-1
114.8 85 203 · 125 since 2021 2026 conflict
Zhaoran Wang 0001
dblp:117/2756-1
107.3 85 143 · 85 since 2021 2026 conflict
Shie Mannor
dblp:20/1669
105.6 111 245 · 66 since 2021 2025 verified
Jianye Hao
dblp:21/7664
103.1 88 288 · 207 since 2021 2026 verified
Doina Precup
dblp:p/DoinaPrecup
99.1 113 218 · 68 since 2021 2026 verified
Marcello Restelli
dblp:64/1011
98.1 93 165 · 92 since 2021 2026 verified
Quanquan Gu
dblp:50/4597
95.0 70 254 · 135 since 2021 2026 conflict
Yaodong Yang 0001
dblp:170/1496-1
94.0 64 126 · 115 since 2021 2026 conflict
Csaba Szepesvári
dblp:62/567
91.2 94 227 · 62 since 2021 2026 verified
Chongjie Zhang
dblp:29/6693
82.3 80 81 · 58 since 2021 2025 corroborated
Wen Sun 0002
dblp:69/1010-2
80.0 75 82 · 52 since 2021 2025 conflict
Alberto Maria Metelli
dblp:209/4941
80.0 65 76 · 65 since 2021 2026 verified
Jakob N. Foerster
dblp:176/5095
79.2 58 97 · 75 since 2021 2026 verified
Weinan Zhang 0001
dblp:28/10261-1
78.4 73 329 · 213 since 2021 2026 conflict
Shimon Whiteson
dblp:42/2548
78.4 78 143 · 44 since 2021 2025 none
Rémi Munos
dblp:69/6815
78.1 105 176 · 39 since 2021 2025 none
Simon S. Du
dblp:176/5602
77.9 68 127 · 88 since 2021 2025 verified
Andreas Krause 0001
dblp:87/1831-1
72.2 59 314 · 125 since 2021 2026 conflict
Jun Wang 0012
dblp:w/JunWang12
68.3 68 258 · 142 since 2021 2026 conflict
Zongzhang Zhang
dblp:90/8724
65.7 63 81 · 58 since 2021 2026 corroborated
Yishay Mansour
dblp:m/YishayMansour
65.6 53 374 · 84 since 2021 2026 verified
Lin Yang 0011
dblp:166/6264
65.2 62 87 · 49 since 2021 2026 verified
Tong Zhang 0001
dblp:07/4227-1
65.1 59 331 · 128 since 2021 2026 conflict
Chelsea Finn
dblp:131/1783
64.0 69 159 · 105 since 2021 2025 verified
Peter Stone 0001
dblp:s/PeterStone
63.7 67 338 · 92 since 2021 2026 conflict
Dacheng Tao
dblp:46/3391
63.4 53 1407 · 608 since 2021 2026 conflict
Amy Zhang 0001
dblp:43/2754-1
62.1 62 50 · 46 since 2021 2025 conflict
Zongqing Lu 0002
dblp:99/965-2
60.4 55 102 · 73 since 2021 2026 conflict
Jan Peters 0001
dblp:p/JanPeters1
59.3 72 285 · 100 since 2021 2026 conflict
Min-hwan Oh
dblp:172/0531
58.7 41 46 · 41 since 2021 2026 conflict
Bo An 0001
dblp:42/6178-1
58.6 56 251 · 144 since 2021 2026 conflict
Aviral Kumar
dblp:202/7961
58.5 45 61 · 52 since 2021 2026 reported
Vaneet Aggarwal
dblp:91/6560
57.4 43 228 · 112 since 2021 2026 verified
Alekh Agarwal
dblp:24/4383
56.6 54 92 · 30 since 2021 2025 corroborated
Huazhe Xu
dblp:164/9006
55.9 51 72 · 63 since 2021 2025 corroborated
Mengdi Wang 0001
dblp:64/10471-1
55.8 57 118 · 85 since 2021 2026 conflict
Nan Jiang 0008
dblp:06/4489-8
55.6 56 59 · 37 since 2021 2025 conflict
page 1 of 754 Next →