Chenxiao Yu

Chenxiao Yu

Ph.D. Student in Computer Science

University of Southern California

Research Interests

Human-Centered AINLPInterpretable and Controllable Foundation ModelsAI Safety and Robustness

About

Chenxiao Yu is a first-year Ph.D. student in Computer Science at the University of Southern California, where he is advised by Prof. Yue Zhao. He also works closely with Prof. Morteza Dehghani. Chenxiao is from Hangzhou, China.

Latest Publications

View All →

Proceedings of the 64th Annual Meeting of the Association for Computational Linguistics · 2026

Defenses Against Prompt Attacks Learn Surface Heuristics

Shawn Li, Chenxiao Yu, Zhiyu Ni, Hao Li, Charith Peris, Chaowei Xiao, Yue Zhao

Current defenses against prompt injection attacks rely on surface-level heuristics rather than robust understanding of attack patterns.

Proceedings of the 43rd International Conference on Machine Learning · 2026

Someone Hid It": Query-Agnostic Black-Box Attacks on LLM-Based Retrieval

Jiate Li, Defu Cao, Li Li, Wei Yang, Yuehan Qin, Chenxiao Yu, Tiannuo Yang, Ryan A. Rossi, Yan Liu, Xiyang Hu, Yue Zhao

A query-agnostic black-box attack framework that manipulates LLM-based retrieval systems without knowledge of user queries.

Proceedings of the AAAI Conference on Artificial Intelligence · 2026

Mitigating Hallucinations in Large Language Models via Causal Reasoning

Yuangang Li, Yiqing Shen, Yi Nian, Jiechao Gao, Ziyi Wang, Chenxiao Yu, Shawn Li, Jie Wang, Xiyang Hu, Yue Zhao

A causal intervention framework that reduces hallucinations in LLMs by identifying and adjusting unstable internal pathways.

News

Jul 2026
Conference

Two papers accepted at ACL 2026 (Main + Findings)! 🎉

Feb 2026
Award

Received the USC Viterbi/Annenberg Fellowship! 🎓

Feb 2026
Award

Received the USC 2026 CCLS Seed Funding Award! 🏆

Dec 2025
Conference

See U Guys @ NeurIPS 2025🎉