← 返回 scale.ai 的题目列表Implement Adversarial Attack using Paper Method
类型:online_judge
Given a paper titled 'Universal and Transferable Adversarial Attacks on Aligned Language Models', read the paper and implement a functionality using the method introduced in the paper. The goal is to attack the GPT-2 model to output certain harmful words. Perform the task in a Google Colab notebook, ensuring the code runs smoothly in the Colab environment.
Example
Input
Example Input 1