publications

Please see my full publication list at google scholar.
* presents equal contribution.

2026

  1. ACL Findings
    InferPilot: Autonomous Inference Attacks Against ML Services With LLM-Based Agents
    Yixin Wu, Rui Wen, Chi Cui, Michael Backes, and Yang Zhang
    In Annual Meeting of the Association for Computational Linguistics (ACL), 2026
  2. ACL Findings
    Peering Behind the Shield: Guardrail Identification in Large Language Models
    Ziqing Yang, Yixin Wu, Rui Wen, Michael Backes, and Yang Zhang
    In Annual Meeting of the Association for Computational Linguistics (ACL), 2026
  3. ACL Findings
    Rethinking Assessments of Prompt Injection Attacks
    Chi Cui, Yixin Wu, Michael Backes, and Yang Zhang
    In Annual Meeting of the Association for Computational Linguistics (ACL), 2026
  4. ECCV
    GEO-Detective: Unveiling Location Privacy Risks in Images with LLM Agents
    Xinyu Zhang, Yixin Wu, Boyang Zhang, Chenhao Lin, Chao Shen, Michael Backes, and Yang Zhang
    In European Conference on Computer Vision (ECCV), 2026
  5. EMNLP Findings
    The Challenge of Identifying the Origin of Black-Box Large Language Models
    Ziqing Yang, Yixin Wu, Yun Shen, Wei Dai, Michael Backes, and Yang Zhang
    In Findings of the Association for Computational Linguistics: EMNLP (EMNLP Findings), 2026
  6. EMNLP Findings
    Beyond Text-Intrinsic Features: An Agentic Framework for Evidence-Based AI-Generated Text Detection
    Chi Cui, Yixin Wu, Michael Backes, and Yang Zhang
    In Findings of the Association for Computational Linguistics: EMNLP (EMNLP Findings), 2026
  7. EMNLP Findings
    GuardrailAgent: An Agentic Framework for Jailbreak Defense
    Tianze Chang, Junjie Chu, Yixin Wu, Michael Backes, and Yang Zhang
    In Findings of the Association for Computational Linguistics: EMNLP (EMNLP Findings), 2026
  8. arxiv
    From Celebrities to Anyone: Characterizing AI Nudification Content, Technology, and Community Dynamics on 4chan
    Chi Cui, Yixin Wu, and Yang Zhang
    CoRR abs/2606.27234, 2026

2025

  1. Usenix Security
    Yixin Wu, Ziqing Yang, Yun Shen, Michael Backes, and Yang Zhang
    In USENIX Security Symposium (USENIX Security), 2025
  2. Usenix Security
    Yixin Wu, Ning Yu, Michael Backes, Yun Shen, and Yang Zhang
    In USENIX Security Symposium (USENIX Security), 2025
  3. Usenix Security
    Xinyue Shen, Yixin Wu, Yiting Qu, Michael Backes, Savvas Zannettou, and Yang Zhang
    In USENIX Security Symposium (USENIX Security), 2025
  4. CCS
    Yiting Qu, Xinyue Shen, Yixin Wu, Michael Backes, Savvas Zannettou, and Yang Zhang
    2025

2024

  1. Usenix Security
    Yixin Wu, Rui Wen, Michael Backes, Pascal Berrang, Mathias Humbert, Yun Shen, and Yang Zhang
    In USENIX Security Symposium (USENIX Security), 2024
  2. CCS
    Yixin Wu, Yun Shen, Michael Backes, and Yang Zhang
    In ACM Conference on Computer and Communications Security (CCS), 2024
  3. PETS
    Yixin Wu, Xinlei He, Pascal Berrang, Mathias Humbert, Michael Backes, Neil Zhenqiang Gong, and Yang Zhang
    In Privacy Enhancing Technologies Symposium (PETS), 2024
  4. EMNLP
    Yihan Ma, Xinyue Shen, Yixin Wu, Boyang Zhang, Michael Backes, and Yang Zhang
    In Empirical Methods in Natural Language Processing (EMNLP), 2024
  5. arxiv
    Xinyue Shen*Yixin Wu*, Michael Backes, and Yang Zhang
    CoRR abs/2405.19103, 2024

2022

  1. arxiv
    Yixin Wu, Ning Yu, Zheng Li, Michael Backes, and Yang Zhang
    CoRR abs/2210.00968, 2022

2021

  1. arxiv
    Xinlei He, Rui Wen, Yixin Wu, Michael Backes, Yun Shen, and Yang Zhang
    CoRR abs/2102.05429, 2021