
ai-security-battle
A Red Team vs. Blue Team Adversarial AI Simulation.

A Red Team vs. Blue Team Adversarial AI Simulation.

Research code for poisoning attacks on the PGM-index, demonstrating how to craft adversarial data to degrade learned index performance.

This project (PoC for now, and part of Shit Bucket) involves face detection, face recognition and adversarial input to protect avatars

Adversarial image perturbation tool that uses SAM segmentation and CLIP models to evade AI-based scam image classifiers for security research.

Research code and experiments for defending tool-integrated LLM agents against adversarial attacks, extending Agent Security Bench with new defense…

Evolutionary LLM jailbreak and guardrail framework that grows a reusable strategy pool via genetic mutation, Markov selection, and online adversarial…


A Python library for Secure and Explainable Machine Learning Documentation available @ https://secml.gitlab.io Follow us on Twitter @…

Fully automatic censorship removal for language models

Test your prompts, agents, and RAGs. Red teaming/pentesting/vulnerability scanning for AI. Compare performance of GPT, Claude, Gemini, DeepSeek, and…

A privacy-first app that strips AI watermarks from content you own.

Modular LLM vulnerability scanner that probes for hallucination, data leakage, prompt injection, jailbreaks, and toxicity using static, dynamic, and…

A Python library for anomaly detection across tabular, time series, graph, text, image, and audio data. 60+ detectors, benchmark-backed ADEngine…

Open-source framework for red-teaming generative AI systems: automate attack prompts, score model responses, and audit behavior to identify security…

🐢 Open-Source Evaluation & Testing library for LLM Agents

Source code about machine learning and security.

reverse engineering Gemini's SynthID detection

A curated list of useful resources that cover Offensive AI.