#1LLM security, prompt injection, model extraction, adversarial AI, and AI red teaming tools.
Kitploit recommended

Meet Eclipse the only jailbreak that moonwalks around ChatGPT 4o.
HopLa Burp Suite Extender plugin - Brings AI capabilities, autocompletion support, and a set of useful payloads to Burp Suite

The OWASP Subtractive Security Top 10 Project is an initiative to identify, document, and promote the highest-impact opportunities for reducing cyber…

Noisegate: a differential privacy gateway that lets an untrusted LLM agent query sensitive data over MCP (Model Context Protocol), with a formal…

Cyber Panel - The hosting control panel for OpenLiteSpeed

Open standard for documenting security-relevant metadata of AI models, including training data provenance, PII risk, known vulnerabilities, and…

Read-only WordPress security scanner for HestiaCP servers. Detects wp2shell compromise indicators (CVE-2026-63030 / CVE-2026-60137) across all hosted…

Collection of CVE(work) on tenserflow binary pwning it

Scan LLM outputs and AI-generated content for data exfiltration signals (EchoLeak, CVE-2025-32711) before they reach users or downstream systems

Human-evaluated benchmark for assessing LLM performance on real-world vulnerability identification, explanation, and remediation across 15+ languages…

Research code and dataset (WSD) for evaluating audio watermarking impact on anti-spoofing, using wav2vec2 XLS-R and knowledge-preserving learning.

This repository provides the official implementation of POISONCRAFT: Practical Poisoning of Retrieval-Augmented Generation for Large Language Models.

Implementation of optimal coupling-based distortion-free text watermarking to embed detectable signals in LLM-generated text while preserving quality.

Ensemble framework for software vulnerability detection and repair using multiple large language models, with consensus analysis and evaluation tools…

[CVPR 2025-ADVML] Official Repository for `Attacking Attention of Foundation Models Effectively Disrupts Downstream Tasks`

Automated framework for hijacking safety reasoning in large reasoning models via simulated reasoning traces and iterative prompt refinement to…

Code Implementation of "Pattern Enhanced Multi-Turn Jailbreaking: Exploiting Structural Vulnerabilities in Large Language Models"

Prohibited Items Segmentation via Occlusion-aware Bilayer Modeling (ICME 2025)