A New Trick Reveals AI Models’ Inner Thoughts
…can lead to personal information leakage, and it enables large-scale reasoning distillation attacks.” Panfilov and colleagues from the University of Tubingen, the Max Planck Institute, the AI safety institute MATS Research…