About me
I am currently a PhD student at CISPA Helmholtz Center for Information Security, advised by Dr. Yang Zhang. I received my Bachelor’s degree from UC Davis (2017).
Research Interests
- Security and Safety of LLMs
- Human-Centered AI System Analysis
Award
- Best Machine Learning and Security Paper in Cybersecurity Award 2025
News
2026.09 I have been invited to serve as a PC member for ACISP 2027.
2026.09 Our paper titled “Can We Trust LLM Security Research? A Reproducibility Study of Large Language Model Security Papers in Tier-1 Security Conferences” got accepted in IEEE S&P 2027!
2026.08 Our paper titled “Real Money, Fake Models: Deceptive Model Claims in Shadow APIs” got accepted in CCS 2026!
2026.07 Our paper titled “Behavior Carries Over: An AI Sandbagging Detection Method Inspired by Honest Inertia” got accepted in EMNLP findings 2026!
2026.07 I have been invited to serve as a PC member for AISec 2026.
2026.04 I am serving as a reviewer for the NeurIPS 2026.
2026.04 Our paper titled “PeerCheck: Enhancing LLM-Generated Academic Reviews Towards Human-Level Quality” got accepted in ACL findings 2026!
2026.01 I am serving as a reviewer for the ICWSM 2026.
2025.09 I have been invited to serve as a PC member for ACISP 2026.
2025.07 Our paper ““Do Anything Now”: Characterizing and Evaluating In-The-Wild Jailbreak Prompts on Large Language Models” won the Best Machine Learning and Security Paper in Cybersecurity Award 2025!
