


Self-Defeating Audits: reproducible lab showing a low-privilege PostgreSQL role reversibly blinding a trigger-based auditor + poisoning attribution…

PoC repository for the blog post CopyEscape: Taking Over Docker Hosts with docker cp

Lifetime AMSI bypass

A toolset to make a system look as if it was the victim of an APT attack



PISmith: Reinforcement Learning-based Red Teaming for Prompt Injection Defenses

Implementation of paper "DeeCLIP: A Robust and Generalizable Transformer-Based Framework for Detecting AI-Generated Images"

The code for ACM MM2024 (Multimodal Unlearnable Examples: Protecting Data against Multimodal Contrastive Learning)

Research code implementing backdoor attack and defense methods for LLMs, including IBSD, SLIP, BeDKD, and BadApex algorithms.

Project Mantis: Hacking Back the AI-Hacker; Prompt Injection as a Defense Against LLM-driven Cyberattacks

Data from a BRAWL Automated Adversary Emulation Exercise

CVPR2023: Unlearnable Clusters: Towards Label-agnostic Unlearnable Examples

A novel adversarial attack on LLM based on the Exponentiated Gradient Descent technique.

AGLS

Code Implementation of "Pattern Enhanced Multi-Turn Jailbreaking: Exploiting Structural Vulnerabilities in Large Language Models"