Send email Copy Email Address

Email

Address

Im Oberen Werk 1
66386 St. Ingbert (Germany)

Awards (selection)

2022: Busy Beaver Award for "Privacy of Machine Learning"

2019: Best paper award at NDSS 

Short Bio

Dr. YAng Zhang is Faculty at CISPA. His research concentrates on trustworthy machine learning (privacy, safety, and security). Moreover, he works on measuring and understanding misinformation and unsafe content like hateful memes on the Internet. Over the years, he has published multiple papers at top venues in computer science, including CCS, NDSS, Oakland, and USENIX Security. His work has received the NDSS 2019 distinguished paper award and the CCS 2022 best paper award runner-up.

CV: Last stations

Since 2020
Faculty at CISPA Helmholtz Center for Information Security
2019 - 2020
Research Group Leader at CISPA Helmholtz Center for Information Security
2017 - 2018
Postdoctoral Researcher - Host: Michael Backes - CISPA, Saarland University
2012 - 2016
Ph.D. in Computer Science at University of Luxembourg, highest honor

Publications by Yang Zhang

Year 2026

Conference / Medium

ACM Conference on Computer and Communications Security (CCS)
BadTV: Unveiling Backdoor Threats in Third-Party Task Vectors

Conference / Medium

Usenix Security Symposium (USENIX-Security)

Conference / Medium

European Conference on Computer Vision (ECCV)
GEO-Detective: Unveiling Location Privacy Risks in Images with LLM Agents

Conference / Medium

Annual Meeting of the Association for Computational Linguistics (ACL)
Reward Yourself: Efficient Self Rewards for Trustworthy Sampling

Conference / Medium

Annual Meeting of the Association for Computational Linguistics (ACL)
PeerCheck: Enhancing LLM-Generated Academic Reviews Towards Human-Level Quality

Conference / Medium

Annual Meeting of the Association for Computational Linguistics (ACL)
Peering Behind the Shield: Guardrail Identification in Large Language Models

Conference / Medium

International Conference on Machine Learning (ICML)
Position: Preparing for AI Systems That Deceive Developers

Conference / Medium

Annual Meeting of the Association for Computational Linguistics (ACL)
Open Schrödinger’s Closed Box: Identifying Retrieval Augmented Generation in API-Accessible Large Language Model Services

Conference / Medium

Annual Meeting of the Association for Computational Linguistics (ACL)

Conference / Medium

IEEE Conference on Computer Vision and Pattern Recognition (CVPR)
When Understanding Becomes a Risk: Authenticity and Safety Risks in the Emerging Image Generation Paradigm