About
I am a Ph.D. candidate in Computing Science at Simon Fraser University, advised by Prof. Ke Wang.
My research is on AI safety and the adversarial robustness of large language models: red-teaming of LLM and computer-use agents, indirect prompt injection, jailbreaking, and safety alignment.
Selected Publications
View All →StepJack: Benchmarking Computer-Use Agent Safety Against Multi-Step Indirect Prompt Injection
Zhuoxin Zhan, Akbar Rafiey, Avery Ma, Leila Pishdad, Layla El Asri
Preprint, arXiv
Benign Prompts Can Jailbreak Large Language Models
Zhuoxin Zhan, Ke Wang, Pulei Xiong, Linyi Li
Under submission
