
T3MP3ST
autonomous red teaming platform; multi-agent offensive-security meta-harness

autonomous red teaming platform; multi-agent offensive-security meta-harness

PHP 8.1.0-dev User-Agentt Backdoor Remote Code Execution (RCE)

LLM-powered agent that autonomously fixes GitHub issues, finds cybersecurity vulnerabilities, and solves CTF challenges using configurable tool-use…

Real-time geospatial OSINT platform aggregating 60+ public telemetry feeds (ADS-B, AIS, satellites, CCTV) into a unified map with server-side recon…

A curated list of GPT agents for cybersecurity

Autonomous Hacking Agent for Red Team

Agent skills for solving CTF challenges - web exploitation, binary pwn, crypto, reverse engineering, forensics, OSINT, and more

A coding-agent skill for multi-phase security audits with independently verified, machine-readable findings

A curated collection of top-tier penetration testing tools and productivity utilities across multiple domains. Join us to explore, contribute, and…

Open-source AI reverse-engineering agent using Ghidra and LLMs to reconstruct and validate C/C++ functions from binaries.

An alignment auditing agent capable of quickly exploring alignment hypothesis

ExploitGym is a large-scale, realistic benchmark built from real-world vulnerabilities designed to evaluate AI agents' ability to develop exploits.

Find zero-days while you sleep. DeepZero is an automated vulnerability research framework that parses, decompiles, and analyzes thousands of Windows…

Evaluation framework for studying LLM agents that automatically generate working exploits from vulnerability reports, bypassing modern security…

A secure* runtime for autonomous AI agents. Policy from plain-English constitutions. (*https://ironcurtain.dev)

An agent to hotpatch the log4j RCE from CVE-2021-44228.

AI / LLM Red Team Field Manual & Consultant’s Handbook