
garak
Modular LLM vulnerability scanner that probes for hallucination, data leakage, prompt injection, jailbreaks, and toxicity using static, dynamic, and…

Modular LLM vulnerability scanner that probes for hallucination, data leakage, prompt injection, jailbreaks, and toxicity using static, dynamic, and…

Fully automatic censorship removal for language models

The agent that grows with you


Framework for auditing machine learning algorithms against adversarial attacks and biases, providing educational tools and an academic paper to…

ESPectre - Motion detection system based on Wi-Fi spectre analysis (CSI), with Home Assistant integration.

The world's most sophisticated street level image geolocation software

This repository includes the source code used in the "Characterization and Detection of Cross-Router Covert Channels" paper.


Research code for red-teaming AI auto-mode monitors, including simulation evals, fuzzing, and monitor implementations for Claude Code and Codex…

YC (S26) | Open Computer History | Continuously record your company computer work, map your workflows, help you find work worth automating, and power…

Benchmark harness measuring where prompt injection defenses fire in tool-using LLM agent pipelines, tracking canary tokens across exposed, persisted,…

Open framework for RL-based prompt injection red teaming, with a shared trainer, curriculum learning, and benchmarks like AgentDojo, InjecAgent, and…

Advanced detection of port scanning, DoS and malware attacks using Machine Learning techniques

The AI that really does things. Any OS. Any Platform. The lobster way. 🦞

An Open-Source Package for Textual Adversarial Attack.

Research implementation for mitigating adaptive prompt injections via on-policy distillation, with training recipes and evaluators for SEP, PISmith,…

AI / LLM Red Team Field Manual & Consultant’s Handbook