The paper introduces AntiSkillBench, an end-to-end benchmark for privacy leakage, attribute disclosure, behavioral impersonation, and defenses in persona-skill pipelines. It contains 7,500 persona-grounded dialogue traces built from 50 behaviorally rich profiles, evaluates three skill-distillation strategies, and tests four online or post-hoc defense configurations. Experiments across three frontier agents suggest that risks persist across backbones and distillation protocols, extending beyond explicit attributes to communication styles and personality traits. Existing defenses show limited and distillation-dependent effectiveness.
No heat snapshots are available in the last 24 hours.