
robustbench
Standardized adversarial robustness benchmark with a public leaderboard and downloadable model zoo for evaluating ML models against Lp attacks and…

Standardized adversarial robustness benchmark with a public leaderboard and downloadable model zoo for evaluating ML models against Lp attacks and…

Unicode encoding attacks with machine learning

Python framework for building LLM workflows as state machines with formal verification via Z3 theorem proving, CTL model checking, and conformal…

Human-evaluated benchmark for assessing LLM performance on real-world vulnerability identification, explanation, and remediation across 15+ languages…

The agent that grows with you

Open-source OCR engine with LSTM neural network models for extracting text from images and scanned documents in 100+ languages via CLI, C/C++…

NVR with realtime local object detection for IP cameras

Test your prompts, agents, and RAGs. Red teaming/pentesting/vulnerability scanning for AI. Compare performance of GPT, Claude, Gemini, DeepSeek, and…

Ready-to-use OCR with 80+ supported languages and all popular writing scripts including Latin, Chinese, Arabic, Devanagari, Cyrillic and etc.

Reverse Engineering: Decompiling Binary Code with Large Language Models

Defeating Google's audio reCaptcha with 85% accuracy.

ESPectre - Motion detection system based on Wi-Fi spectre analysis (CSI), with Home Assistant integration.

🥂 Gracefully face hCaptcha challenge with multimodal large language model.

AI-powered threat intelligence platform for automated CVE/ransomware monitoring, domain surveillance, data leak detection, and incident response with…

Automated behavioral evaluation framework for LLMs that generates diverse test scenarios to probe for sycophancy, bias, and other safety-relevant…

OGhidra bridges Large Language Models (LLMs) via Ollama with the Ghidra reverse engineering platform, enabling AI-driven binary analysis through…

Automating Host Exploitation with AI

Train, evaluate, and explore neural networks with built-in adversarial robustness tools, including PGD attacks, adversarial training, and robust…