Co-occurring keywords
Papers
Improving Authorship Privacy: Adaptive Obfuscation with the Dynamic Selection of Techniques
ACL 2024
Deconstructing Classifiers: Towards A Data Reconstruction Attack Against Text Classification Models
ACL 2024
CycleAlign: Iterative Distillation from Black-box LLM to White-box Models for Better Human Alignment
ACL 2024
TLCR: Token-Level Continuous Reward for Fine-grained Reinforcement Learning from Human Feedback
ACL 2024