AAAI 2026

•

January 24, 2026

•

Singapore, Singapore

Please log in to leave a comment

Downloads

SlidesPaperTranscript English (automatic)

Next from AAAI 2026

Uncovering and Aligning Anomalous Attention Heads to Defend Against NLP Backdoor Attacks
poster

Uncovering and Aligning Anomalous Attention Heads to Defend Against NLP Backdoor Attacks

AAAI 2026

Haotian Jin
+3
Haihui Fan and 5 other authors

24 January 2026

Similar lecture

RLHFPoison: Reward Poisoning Attack for Reinforcement Learning with Human Feedback in Large Language Models
poster

RLHFPoison: Reward Poisoning Attack for Reinforcement Learning with Human Feedback in Large Language Models

ACL 2024

Yevgeniy Vorobeychik
+2
Jiongxiao Wang and 4 other authors

12 August 2024