EDBT 2026 Demo / reviewers in the wild / expert
Xinge Peng
dblp:352/5837
· DBLP profile ↗
1ranked-venue papers
0as first author
1since 2021 · last 2025
0009-0006-3820-8810ORCID · reported
Domains — the database's venue-derived domains; a paper can count in several
Graphics, computer vision, multimedia, augmented reality and games · 1 · 1 since 2021
Expertise — from the expertise taxonomy: the topics of the expert's papers under the CCF categories. A weight counts papers with recency: 1 for a paper about the topic, 0.3 when the topic is its context, halved every five years.
| Artificial intelligence
1 paper |
Video understanding and tracking · 100% |
Topics — the 2 heaviest of 2, each with the papers that count most for it
| Topic | Weight | Papers | Last | Evidence papers |
|---|---|---|---|---|
Computer vision › Video understanding and tracking › human action analysis › action understanding › action counting
repetitive action counting |
0.9 | 1 | 2025 | Repetitive Action Counting With Hybrid Temporal Relation Modeling · IEEE Trans. Multim. 2025 |
Computer vision › Video understanding and tracking › temporal modeling
temporal relation modeling |
0.3 | 1 | 2025 | Repetitive Action Counting With Hybrid Temporal Relation Modeling · IEEE Trans. Multim. 2025 |
Methods — techniques the papers use, named apart from their topics
temporal self-similarity matrix · 0.9self-attention · 0.9multi-scale matrix fusion · 0.9dual softmax · 0.9
| Year | Publication | Venue | Position |
|---|---|---|---|
| 2025 | Repetitive Action Counting With Hybrid Temporal Relation ModelingabstractRepetitive Action Counting (RAC) aims to count the number of repetitive actions occurring in videos. In the real world, repetitive actions have great diversity and bring numerous challenges (e.g., viewpoint changes, non-uniform periods, and action interruptions). Existing methods based on the temporal self-similarity matrix (TSSM) for RAC are trapped in the bottleneck of insufficient capturing action periods when applied to complicated daily videos. To tackle this issue, we propose a novel method named Hybrid Temporal Relation Modeling Network (HTRM-Net) to build diverse TSSM for RAC. The HTRM-Net mainly consists of three key components: bi-modal temporal self-similarity matrix modeling, random matrix dropping, and local temporal context modeling. Specifically, we construct temporal self-similarity matrices by bi-modal (self-attention and dual-softmax) operations, yielding diverse matrix representations from the combination of row-wise and column-wise correlations. To further enhance matrix representations, we propose incorporating a random matrix dropping module to guide channel-wise learning of the matrix explicitly. After that, we inject the local temporal context of video frames and the learned matrix into temporal correlation modeling, which can make the model robust enough to cope with error-prone situations, such as action interruption. Finally, a multi-scale matrix fusion module is designed to aggregate temporal correlations adaptively in multi-scale matrices. Extensive experiments across intra- and cross-datasets demonstrate that the proposed method not only outperforms current state-of-the-art methods and but also exhibits robust capabilities in accurately counting repetitive actions in unseen action categories. Notably, our method surpasses the classical TransRAC method by 20.04% in MAE and 22.76% in OBO. Kun Li 0008, Xinge Peng, Dan Guo 0001, Xun Yang 0001, Meng Wang 0001 |
IEEE Trans. Multim. | 2 |