Conference Intel
Best Talks Tools Rate My Talk About
Home / IEEE Symposium on Security and Privacy 2025 / Preference Poisoning Attacks on Reward Model Learning

Preference Poisoning Attacks on Reward Model Learning

Junlin Wu, Jiongxiao Wang, Chaowei Xiao, Chenguang Wang, Ning Zhang, Yevgeniy Vorobeychik

IEEE Symposium on Security and Privacy 2025 · Day 2 · ML Attacks

← Papercraft: Lattice-based Verifiable Delay Function Implemented · Query Provenance Analysis: Efficient and Robust Defense against Query-based… →

All talks from IEEE Symposium on Security and Privacy 2025

[ Conference Intel ]

Community-driven deep-dive coverage of leading security conferences.

Subscribe via RSS — new talk reviews as they publish.