Ziyue Wang 0002

dblp:137/0610-2 · DBLP profile ↗
← Back
7ranked-venue papers
3as first author
7since 2021 · last 2026
0009-0004-1433-0681ORCID · conflict

Domains — the database's venue-derived domains; a paper can count in several

Artificial intelligence and machine learning · 7 · 3 first-author · 7 since 2021Graphics, computer vision, multimedia, augmented reality and games · 1 · 1 first-author · 1 since 2021

Expertise — from the expertise taxonomy: the topics of the expert's papers under the CCF categories. A weight counts papers with recency: 1 for a paper about the topic, 0.3 when the topic is its context, halved every five years.

Artificial intelligence
6 papers
Vision and language · 64% Reinforcement learning · 11% Robot navigation and mapping · 9%

Topics — the 9 heaviest of 10, each with the papers that count most for it

TopicWeightPapersLastEvidence papers
Computer vision › Vision and language › vision-language model
multimodal large language model
1.622025
ActiView: Evaluating Active Perception Ability for Multimodal Large Language Models · ACL (1) 2025
Model Composition for Multimodal Large Language Models · ACL (1) 2024
Computer vision › Vision and language › vision-language model › multimodal large language model
multimodal large language model evaluation
1.622025
How Do Multimodal Large Language Models Handle Complex Multimodal Reasoning? Placing Them in an Extensible Escape Game · ICCV 2025
CODIS: Benchmarking Context-dependent Visual Comprehension for Multimodal Large Language Models · ACL (1) 2024
Machine learning › Reinforcement learning
reward design
1.012026
MUSEG: Reinforcing Video Temporal Understanding via Timestamp-Aware Multi-Segment Grounding · ACL (1) 2026
Computer vision › Vision and language
temporal grounding
1.012026
MUSEG: Reinforcing Video Temporal Understanding via Timestamp-Aware Multi-Segment Grounding · ACL (1) 2026
Robotics › Robot navigation and mapping
active perception
0.912025
ActiView: Evaluating Active Perception Ability for Multimodal Large Language Models · ACL (1) 2025
Computer vision › Vision and language › vision-language model › multimodal large language model
multimodal large language model reasoning
0.912025
How Do Multimodal Large Language Models Handle Complex Multimodal Reasoning? Placing Them in an Extensible Escape Game · ICCV 2025
Machine learning › Efficient and distributed learning
model composition
0.812024
Model Composition for Multimodal Large Language Models · ACL (1) 2024
Computer vision › Vision and language
multimodal understanding
0.812024
Browse and Concentrate: Comprehending Multimodal Content via Prior-LLM Context Fusion · ACL (1) 2024
Computer vision › Segmentation and scene understanding › semantic segmentation
context aggregation
0.212024
Browse and Concentrate: Comprehending Multimodal Content via Prior-LLM Context Fusion · ACL (1) 2024

Methods — techniques the papers use, named apart from their topics

video-language model · 1.0reinforcement learning · 1.0multimodal large language model evaluation · 0.9escape game benchmark · 0.9prior-LLM context fusion · 0.8multimodal learning · 0.8model merging · 0.8benchmark construction · 0.8
YearPublicationVenuePosition
2026 MUSEG: Reinforcing Video Temporal Understanding via Timestamp-Aware Multi-Segment Grounding
abstract
Fuwen Luo, Shengfeng Lou, Chi Chen, Ziyue Wang, Chenliang Li, Weizhou Shen, Jiyue Guo, Peng Li, Ming Yan, Ji Zhang, Fei Huang, Yang Liu. Proceedings of the 64th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers). 2026.
Fuwen Luo, Shengfeng Lou, Chi Chen 0005, Ziyue Wang 0002, Chenliang Li 0003, Weizhou Shen, Jiyue Guo, Peng Li 0030, Ming Yan 0008, Ji Zhang 0011, Fei Huang 0002, Yang Liu 0005
ACL (1)4
2025 ActiView: Evaluating Active Perception Ability for Multimodal Large Language Models
abstract
Ziyue Wang, Chi Chen, Fuwen Luo, Yurui Dong, Yuanchi Zhang, Yuzhuang Xu, Xiaolong Wang, Peng Li, Yang Liu. Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers). 2025.
Ziyue Wang 0002, Chi Chen 0005, Fuwen Luo, Yurui Dong 0001, Yuanchi Zhang, Yuzhuang Xu, Xiaolong Wang 0014, Peng Li 0030, Yang Liu 0005
ACL (1)1
2025 How Do Multimodal Large Language Models Handle Complex Multimodal Reasoning? Placing Them in an Extensible Escape Game
Ziyue Wang 0002, Yurui Dong 0001, Fuwen Luo, Minyuan Ruan, Zhili Cheng, Chi Chen 0005, Peng Li 0030, Yang Liu 0005
ICCV1
2024 Model Composition for Multimodal Large Language Models
abstract
Chi Chen, Yiyang Du, Zheng Fang, Ziyue Wang, Fuwen Luo, Peng Li, Ming Yan, Ji Zhang, Fei Huang, Maosong Sun, Yang Liu. Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers). 2024.
Chi Chen 0005, Yiyang Du, Ziyue Wang 0002, Fuwen Luo, Peng Li 0030, Ming Yan 0008, Ji Zhang 0011, Fei Huang 0002, Maosong Sun 0001, Yang Liu 0005
ACL (1)4
2024 CODIS: Benchmarking Context-dependent Visual Comprehension for Multimodal Large Language Models
abstract
Fuwen Luo, Chi Chen, Zihao Wan, Zhaolu Kang, Qidong Yan, Yingjie Li, Xiaolong Wang, Siyu Wang, Ziyue Wang, Xiaoyue Mi, Peng Li, Ning Ma, Maosong Sun, Yang Liu. Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers). 2024.
Fuwen Luo, Chi Chen 0005, Zihao Wan, Zhaolu Kang, Qidong Yan, Yingjie Li 0009, Xiaolong Wang 0014, Ziyue Wang 0002, Xiaoyue Mi, Peng Li 0030, Maosong Sun 0001, Yang Liu 0005
ACL (1)9
2024 Browse and Concentrate: Comprehending Multimodal Content via Prior-LLM Context Fusion
abstract
Ziyue Wang, Chi Chen, Yiqi Zhu, Fuwen Luo, Peng Li, Ming Yan, Ji Zhang, Fei Huang, Maosong Sun, Yang Liu. Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers). 2024.
Ziyue Wang 0002, Chi Chen 0005, Yiqi Zhu, Fuwen Luo, Peng Li 0030, Ming Yan 0008, Ji Zhang 0011, Fei Huang 0002, Maosong Sun 0001, Yang Liu 0005
ACL (1)1
2023 TiBERT: A Non-autoregressive Pre-trained Model for Text Editing
Baoxin Wang, Ziyue Wang 0002, Wanxiang Che, Dayong Wu, Shijin Wang 0001
NLPCC (3)2