Preprint
Aug 2026
GRASP: Reinforcing Language Model Anonymizers with Group Relative Policy Optimization
A single small model acts as anonymizer, adversary, and utility judge, trained against a self-generated reward that hides attributes while preserving meaning, with a design that guards against reward hacking.
Sajjad Ghiasvand, Nader Sehatbakhsh
· 0 citations