VLDB 2026 Research / reviewers in the wild / expert
Jiahe Guo
dblp:376/3062
· DBLP profile ↗
11ranked-venue papers
2as first author
11since 2021 · last 2026
—ORCID · conflict
Domains — the database's venue-derived domains; a paper can count in several
Artificial intelligence and machine learning · 9 · 1 first-author · 9 since 2021Computer networks · 1 · 1 first-author · 1 since 2021Graphics, computer vision, multimedia, augmented reality and games · 1 · 1 since 2021
Expertise — from the expertise taxonomy: the topics of the expert's papers under the CCF categories. A weight counts papers with recency: 1 for a paper about the topic, 0.3 when the topic is its context, halved every five years.
| Artificial intelligence
6 papers |
Language models and text generation · 36% Trustworthy machine learning · 25% Transfer learning and domain adaptation · 15% | |
| Network and information security
1 paper |
Security and privacy of machine learning · 100% |
Topics — the 10 heaviest of 13, each with the papers that count most for it
| Topic | Weight | Papers | Last | Evidence papers |
|---|---|---|---|---|
Machine learning › Transfer learning and domain adaptation
fine-tuning |
1.7 | 2 | 2025 | MPO: Multilingual Safety Alignment via Reward Gap Optimization · ACL (1) 2025 Beware of Your Po! Measuring and Mitigating AI Safety Risks in Role-Play Fine-Tuning of LLMs · ACL (1) 2025 |
Natural language and speech › Language models and text generation
chain-of-thought reasoning |
1.0 | 1 | 2026 | Trade-offs in Large Reasoning Models: An Empirical Analysis of Deliberative and Adaptive Reasoning over Foundational Capabilities · AAAI 2026 |
Natural language and speech › Question answering and dialogue systems › open-domain dialogue
emotional support conversation |
1.0 | 1 | 2026 | TEA-Bench: A Systematic Benchmarking of Tool-enhanced Emotional Support Dialogue Agent · ACL (1) 2026 |
Natural language and speech › Language models and text generation
hallucination mitigation |
1.0 | 1 | 2026 | TEA-Bench: A Systematic Benchmarking of Tool-enhanced Emotional Support Dialogue Agent · ACL (1) 2026 |
Natural language and speech › Language models and text generation › large language model
large reasoning model |
1.0 | 1 | 2026 | Trade-offs in Large Reasoning Models: An Empirical Analysis of Deliberative and Adaptive Reasoning over Foundational Capabilities · AAAI 2026 |
Machine learning › Representation and self-supervised learning › representation learning
disentangled representation learning |
0.9 | 1 | 2025 | When Less Language is More: Language-Reasoning Disentanglement Makes LLMs Better Multilingual Reasoners · NeurIPS 2025 |
Natural language and speech › Language models and text generation › large language model reasoning
multilingual reasoning |
0.9 | 1 | 2025 | When Less Language is More: Language-Reasoning Disentanglement Makes LLMs Better Multilingual Reasoners · NeurIPS 2025 |
Machine learning › Trustworthy machine learning › AI safety
safety alignment |
0.9 | 1 | 2025 | Beware of Your Po! Measuring and Mitigating AI Safety Risks in Role-Play Fine-Tuning of LLMs · ACL (1) 2025 |
Security and privacy of machine learning › large language model safety
jailbreak defense |
0.9 | 1 | 2025 | AdaSteer: Your Aligned LLM is Inherently an Adaptive Jailbreak Defender · EMNLP 2025 |
Natural language and speech › Language models and text generation › alignment
aligned large language models |
0.3 | 1 | 2025 | AdaSteer: Your Aligned LLM is Inherently an Adaptive Jailbreak Defender · EMNLP 2025 |
Methods — techniques the papers use, named apart from their topics
representation steering · 1.7adversarial prompting · 1.7tool augmentation · 1.0supervised fine-tuning · 1.0empirical analysis · 1.0adaptive reasoning · 1.0safety risk measurement · 0.9reward gap optimization · 0.9mitigation · 0.9causal intervention · 0.9
| Year | Publication | Venue | Position |
|---|---|---|---|
| 2026 | Trade-offs in Large Reasoning Models: An Empirical Analysis of Deliberative and Adaptive Reasoning over Foundational CapabilitiesabstractRecent advancements in Large Reasoning Models (LRMs), such as OpenAI's o1/o3 and DeepSeek-R1, have demonstrated remarkable performance in specialized reasoning tasks through human-like deliberative thinking and long chain-of-thought reasoning. However, our systematic evaluation across various model families (DeepSeek, Qwen, and LLaMA) and scales (7B to 32B) reveals that acquiring these deliberative reasoning capabilities significantly reduces the foundational capabilities of LRMs, including notable declines in helpfulness and harmlessness, alongside substantially increased inference costs. Importantly, we demonstrate that adaptive reasoning---employing modes like Zero-Thinking, Less-Thinking, and Summary-Thinking---can effectively alleviate these drawbacks. Our empirical insights underline the critical need for developing more versatile LRMs capable of dynamically allocating inference-time compute according to specific task characteristics. Weixiang Zhao, Xingyu Sui, Jiahe Guo, Yulin Hu, Yang Deng 0002, Xuda Zhi, Yongbo Huang, Wanxiang Che, Ting Liu 0001, Bing Qin 0001 |
AAAI | 3 |
| 2026 | When Personalization Legitimizes Risks: Uncovering Safety Vulnerabilities in Personalized Dialogue AgentsabstractJiahe Guo, Xiangran Guo, Yulin Hu, Zimo Long, Xingyu Sui, Xuda Zhi, Yongbo Huang, Hao He, Weixiang Zhao, Yanyan Zhao, Bing Qin. Proceedings of the 64th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers). 2026. Jiahe Guo, Xiangran Guo, Yulin Hu, Zimo Long, Xingyu Sui, Xuda Zhi, Yongbo Huang, Weixiang Zhao, Bing Qin 0001 |
ACL (1) | 1 |
| 2026 | TEA-Bench: A Systematic Benchmarking of Tool-enhanced Emotional Support Dialogue AgentabstractEmotional Support Conversation requires not only affective expression but also grounded instrumental support to provide trustworthy guidance.However, existing ESC systems and benchmarks largely focus on affective support in text-only settings, overlooking how external tools can enable factual grounding and reduce hallucination in multi-turn emotional support.We introduce TEA-Bench, the first interactive benchmark for evaluating tool-augmented agents in ESC, featuring realistic emotional scenarios, an MCP-style tool environment, and process-level metrics that jointly assess the quality and factual grounding of emotional support.Experiments on nine LLMs show that tool augmentation generally improves emotional support quality and reduces hallucination, but the gains are strongly capacity-dependent: stronger models use tools more selectively and effectively, while weaker models benefit only marginally.We further release TEA-Dialog, a dataset of toolenhanced ESC dialogues, and find that supervised fine-tuning improves in-distribution support but generalizes poorly.Our results underscore the importance of tool use in building reliable emotional support agents. 1 Xingyu Sui, Yulin Hu, Jiahe Guo, Weixiang Zhao, Bing Qin 0001 |
ACL (1) | 4 |
| 2025 | Beware of Your Po! Measuring and Mitigating AI Safety Risks in Role-Play Fine-Tuning of LLMsabstractWeixiang Zhao, Yulin Hu, Yang Deng, Jiahe Guo, Xingyu Sui, Xinyang Han, An Zhang, Yanyan Zhao, Bing Qin, Tat-Seng Chua, Ting Liu. Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers). 2025. Weixiang Zhao, Yulin Hu, Yang Deng 0002, Jiahe Guo, Xingyu Sui, An Zhang 0003, Bing Qin 0001, Tat-Seng Chua, Ting Liu 0001 |
ACL (1) | 4 |
| 2025 | MPO: Multilingual Safety Alignment via Reward Gap OptimizationabstractWeixiang Zhao, Yulin Hu, Yang Deng, Tongtong Wu, Wenxuan Zhang, Jiahe Guo, An Zhang, Yanyan Zhao, Bing Qin, Tat-Seng Chua, Ting Liu. Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers). 2025. Weixiang Zhao, Yulin Hu, Yang Deng 0002, Tongtong Wu, Wenxuan Zhang 0001, Jiahe Guo, An Zhang 0003, Bing Qin 0001, Tat-Seng Chua, Ting Liu 0001 |
ACL (1) | 6 |
| 2025 | AdaSteer: Your Aligned LLM is Inherently an Adaptive Jailbreak DefenderabstractWeixiang Zhao, Jiahe Guo, Yulin Hu, Yang Deng, An Zhang, Xingyu Sui, Xinyang Han, Yanyan Zhao, Bing Qin, Tat-Seng Chua, Ting Liu. Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing. 2025. Weixiang Zhao, Jiahe Guo, Yulin Hu, Yang Deng 0002, An Zhang 0003, Xingyu Sui, Bing Qin 0001, Tat-Seng Chua, Ting Liu 0001 |
EMNLP | 2 |
| 2025 | Real-Time Anomaly Detection and Completion in Data Streams Based on RRCF and RTHaLRTC Tensor Completion
Tong Liang, Binghua Li 0001, Ziqing Chang, Chao Li 0013, Jiahe Guo, Jordi Solé i Casals, Yasuhiro Kushihashi, Ryutaro Himeno, Zhe Sun 0009 |
ICONIP (3) | 6 |
| 2025 | When Less Language is More: Language-Reasoning Disentanglement Makes LLMs Better Multilingual ReasonersabstractMultilingual reasoning remains a significant challenge for large language models (LLMs), with performance disproportionately favoring high-resource languages. Drawing inspiration from cognitive neuroscience, which suggests that human reasoning functions largely independently of language processing, we hypothesize that LLMs similarly encode reasoning and language as separable components that can be disentangled to enhance multilingual reasoning. To evaluate this, we perform a causal intervention by ablating language-specific representations at inference time. Experiments on 10 open-weight LLMs spanning 11 typologically diverse languages show that this language-specific ablation consistently boosts multilingual reasoning performance. Layer-wise analyses further confirm that language and reasoning representations can be effectively disentangled throughout the model, yielding improved multilingual reasoning capabilities, while preserving top-layer language features remains essential for maintaining linguistic fidelity. Compared to post-training methods such as supervised fine-tuning or reinforcement learning, our training-free language-reasoning disentanglement achieves comparable or superior results with minimal computational overhead. These findings shed light on the internal mechanisms underlying multilingual reasoning in LLMs and suggest a lightweight and interpretable strategy for improving cross-lingual generalization. Weixiang Zhao, Jiahe Guo, Yang Deng 0002, Tongtong Wu, Wenxuan Zhang 0001, Yulin Hu, Xingyu Sui, Wanxiang Che, Bing Qin 0001, Tat-Seng Chua, Ting Liu 0001 |
NeurIPS | 2 |
| 2025 | Teaching Language Models to Evolve with Users: Dynamic Profile Modeling for Personalized AlignmentabstractPersonalized alignment is essential for enabling large language models (LLMs) to engage effectively in user-centric dialogue. While recent prompt-based and offline optimization methods offer preliminary solutions, they fall short in cold-start scenarios and long-term personalization due to their inherently static and shallow designs. In this work, we introduce the Reinforcement Learning for Personalized Alignment (RLPA) framework, in which an LLM interacts with a simulated user model to iteratively infer and refine user profiles through dialogue. The training process is guided by a dual-level reward structure: the Profile Reward encourages accurate construction of user representations, while the Response Reward incentivizes generation of responses consistent with the inferred profile. We instantiate RLPA by fine-tuning Qwen-2.5-3B-Instruct, resulting in Qwen-RLPA, which achieves state-of-the-art performance in personalized dialogue. Empirical evaluations demonstrate that Qwen-RLPA consistently outperforms prompting and offline fine-tuning baselines, and even surpasses advanced commercial models such as Claude-3.5 and GPT-4o. Further analysis highlights Qwen-RLPA's robustness in reconciling conflicting user preferences, sustaining long-term personalization and delivering more efficient inference compared to recent reasoning-focused LLMs. These results emphasize the potential of dynamic profile inference as a more effective paradigm for building personalized dialogue systems. Weixiang Zhao, Xingyu Sui, Yulin Hu, Jiahe Guo, Haixiao Liu, Biye Li, Bing Qin 0001, Ting Liu 0001 |
NeurIPS | 4 |
| 2025 | Movable Antenna Aided ISAC with Non-Orthogonal Multiple Access: Joint Power Allocation, Beamforming and Antenna Position DesignabstractThis paper investigates a movable antenna (MA)-aided integrated sensing and communication (ISAC) system using non-orthogonal multiple access (NOMA). A base station (BS) configured with a two-dimensional MA array simultaneously serves multiple communication users and sensing multiple targets. To enhance system capacity, superimposed symbols are transmitted to communication users, with successive interference cancellation (SIC) employed for signal decoding. Our objective is to maximize the total illumination power at the targets while satisfying the minimum signal-to-interference-plus-noise ratio (SINR) requirements for communication users. To achieve this goal, we propose an alternating optimization (AO)-based algorithm that jointly optimizes the transmit power allocation, beamforming, sensing covariance matrix, and antenna positions. Numerical results show that the MA system achieves significant improvement in illumination power compared to fixed-position antennas (FPAs), with particularly significant gains under high SINR requirements. Wanting Lyu, Baojuan Liu, Yue Xiu 0001, Zhongpei Zhang, Jiahe Guo, Chadi Assi, Chau Yuen |
PIMRC | 5 |
| 2025 | Flexible Cylindrical Arrays With Movable Antennas for MISO System: Beamforming and Position OptimizationabstractAs wireless communication advances toward the 6G era, the demand for ultra-reliable, high-speed, and ubiquitous connectivity is driving the exploration of new degrees-of-freedom (DoFs) in communication systems. Among the key enabling technologies, Movable Antennas (MAs) integrated into Flexible Cylindrical Arrays (FCLA) have shown great potential in optimizing wireless communication by providing spatial flexibility. This paper proposes an innovative optimization framework that leverages the dynamic mobility of FCLAs to improve communication rates and overall system performance. By employing Fractional Programming (FP) for alternating optimization of beamforming and antenna positions, the system enhances throughput and resource utilization. Additionally, a novel Constrained Grid Search-Based Adaptive Moment Estimation Algorithm (CGS-Adam) is introduced to optimize antenna positions while adhering to antenna spacing constraints. Extensive simulations validate that the proposed system, utilizing movable antennas, significantly outperforms traditional fixed antenna optimization, achieving up to a 31% performance gain in general scenarios. The integration of FCLAs in wireless networks represents a promising solution for future 6G systems, offering improved coverage, energy efficiency, and flexibility. Jiahe Guo, Songjie Yang, Jiapan Yang, Junfeng Deng, Zhongpei Zhang, Chau Yuen |
IEEE Internet Things J. | 1 |