
adversarial-robustness-toolbox
Python library for adversarial machine learning security, enabling red and blue teams to run evasion, poisoning, extraction, and inference attacks…

Python library for adversarial machine learning security, enabling red and blue teams to run evasion, poisoning, extraction, and inference attacks…

Set of tools to assess and improve LLM security.

Tools that trigger False Positive AV alerts

An autonomous red-teaming engine for LLMs. RedThread manages the full security lifecycle: generating adversarial attacks, executing precision…

Detects LLM context-leakage attacks by training lightweight behavior probes on log-probabilities, with vLLM offline/server detection pipelines.

A structured knowledge base covering AI security fundamentals, threat modeling, red team offensive techniques, and blue team defenses, including LLM…

An information security preparedness tool to do adversarial simulation.

Algorithms for outlier, adversarial and drift detection

Train, evaluate, and explore neural networks with built-in adversarial robustness tools, including PGD attacks, adversarial training, and robust…

Protection against Model Serialization Attacks

An adversarial example library for constructing attacks, building defenses, and benchmarking both

A Python library for anomaly detection across tabular, time series, graph, text, image, and audio data. 60+ detectors, benchmark-backed ADEngine…

Detection-aware BloodHound attack-path scoring - the quietest route to your objective, calibrated across five detection tiers…

The Security Toolkit for LLM Interactions

Multi-layered prompt injection detector for AI applications using heuristics, LLM-based analysis, vectorDB attack signatures, and canary token leak…

Purple Team Exercise Framework

Open-source AI security platform providing perimeter defense for LLMs and AI agents through swarm analysis, policy enforcement, adversarial testing,…

Reverse Shell Detection with Machine Learning