Skip to content
KitploitKITPLOIT
ToolsBlog
Log in
Submit
ToolsBlog
Submit

Hacking, PenTest, and Cybersecurity Tools for Your Security Arsenal!

Kitploit is a directory of hacking, cybersecurity, and pentesting tools. Discover the latest project updates to find vulnerabilities, analyze systems, automate testing, and strengthen your security.

··Feeds·Contact·Privacy·© 2026 Kitploit

Tool Directory

Categories

View all categories
Loading categories
Exponentiated-Gradient-Descent-LLM-Attack — A novel adversarial attack on LLM based on the Exponentiated Gradient Descent technique. | Kitploit
Tools/GitHubGitHub/sbamit/exponentiated-gradient-descent-llm-attack
Machine LearningPapers & ResearchAI SecurityAdversarial Attack
GitHubsbamit/exponentiated-gradient-descent-llm-attack

Exponentiated-Gradient-Descent-LLM-Attack

A novel adversarial attack on LLM based on the Exponentiated Gradient Descent technique.

View Repository

Most Popular

View all →

Discover the most used tools by our community.

Explore all tools

Browse our collection of tools

View all tools →
Share
41149 months agoNot yet reviewed

Change Readme File.

This is a Project that explores the Exponentiated Gradient Descent optimizaiton method to produce adversarial suffix to attack algined Large Language Models. The method is shown to be effective on Llama-2 chat model with 7 Billion parameters.

To run pgd script on a number of behaviors(i.e., 20), execute the following command in the shell. python run_pgd.py
--input_file "/home/samuel/research/llmattacks/llm-attacks/data/advbench/harmful_behaviors.csv"
--output_file "./JSON_Files/PGD_AdvBench_Llama2.jsonl"
--model "Llama2"
--dataset_name "AdvBench"
--num_behaviors 20

Next target is to build similar pipeline for the EGD with Adam optim script.

Download Tool