
PurpleLlama
Set of tools to assess and improve LLM security.
adversarial-attackai-securitycode-analysis+4
4.4k7 days ago

Set of tools to assess and improve LLM security.

Project Mantis: Hacking Back the AI-Hacker; Prompt Injection as a Defense Against LLM-driven Cyberattacks

A comprehensive set of fairness metrics for datasets and machine learning models, explanations for these metrics, and algorithms to mitigate bias in…

Lifetime AMSI bypass