EDBT 2026 Demo / reviewers in the wild / expert
Yangchen Dong
dblp:238/0587
· DBLP profile ↗
3ranked-venue papers
0as first author
3since 2021 · last 2024
—ORCID · none
Domains — the database's venue-derived domains; a paper can count in several
Systems, architecture and hardware · 3 · 3 since 2021
| Year | Publication | Venue | Position |
|---|---|---|---|
| 2024 | Reinforcement Learning-powered Effectiveness and Efficiency Few-shot Jailbreaking Attack LLMsabstractThe widespread use of large language models (LLMs) has brought about security risks, including biases, discrimination, and ethical concerns. Reinforcement Learning from Human Feedback (RLHF), as a method to improve model security, still faces challenges such as objective management and misaligned generalization, leading to the emergence of jailbreak attacks. Existing methods implement jailbreak attacks by optimizing adversarial prompts or leveraging the in-context learning capabilities of LLMs, but they are limited in terms of efficiency and scalability. This paper proposes a reinforcement learning-based few-shot example selection method to enhance the effectiveness and efficiency of these attacks. The proposed method extends the GPT-2 architecture with an example selection module and employs strategies such as experience replay and entropy penalty to accelerate convergence and avoid local optima. Experimental results demonstrate that, compared to existing methods, this approach achieves a 100% increase in attack success rate on Vicuna-7B and a 2.4-second reduction in the time cost per harmful instruction generation on GPT-3.5. Xuehai Tang, Zhongjiang Yao, Jie Wen 0007, Yangchen Dong, Jizhong Han, Songlin Hu 0001 |
ISPA | 4 |
| 2023 | A Multi-source Domain Adaption Approach to Minority Disk Failure Prediction
Wang Wang, Xuehai Tang, Biyu Zhou, Yangchen Dong, Yuanhang Feng, Jizhong Han, Songlin Hu 0001 |
ICA3PP (2) | 4 |
| 2021 | Fed-Tra: Improving Accuracy of Deep Learning Model on Non-iid in Federated Learning
Wenjie Xiao, Xuehai Tang, Biyu Zhou, Wang Wang, Yangchen Dong, Liangjun Zang, Jizhong Han, Songlin Hu 0001 |
ICA3PP (1) | 5 |