Zhuoxin Zhan

Zhuoxin Zhan

Ph.D. Candidate

Simon Fraser University

Research Interests

LLM & Agent Safety
Indirect Prompt Injection
Jailbreaking & Safety Alignment
Adversarial Robustness

About

I am a Ph.D. candidate in Computing Science at Simon Fraser University, advised by Prof. Ke Wang.

My research is on AI safety and the adversarial robustness of large language models: red-teaming of LLM and computer-use agents, indirect prompt injection, jailbreaking, and safety alignment.

Selected Publications

View All

StepJack: Benchmarking Computer-Use Agent Safety Against Multi-Step Indirect Prompt Injection

Zhuoxin Zhan, Akbar Rafiey, Avery Ma, Leila Pishdad, Layla El Asri

Preprint, arXiv

Benign Prompts Can Jailbreak Large Language Models

Zhuoxin Zhan, Ke Wang, Pulei Xiong, Linyi Li

Under submission

Accelerating Adversarial Training on Under-Utilized GPU

Zhuoxin Zhan, Ke Wang, Pulei Xiong

Proceedings of the 34th International Joint Conference on Artificial Intelligence (IJCAI)