Imperio

(★ 44)

[IJCAI 2024] Imperio is an LLM-powered backdoor attack. It allows the adversary to issue language-guided instructions to control the victim model's prediction for arbitrary targets.

Imperio Latest Version Download

Download Latest Version (.zip)
// repository documentation