
AI_HACKING
Collection of CVE(work) on tenserflow binary pwning it

Collection of CVE(work) on tenserflow binary pwning it

This repository provides the official implementation of POISONCRAFT: Practical Poisoning of Retrieval-Augmented Generation for Large Language Models.

A modular framework for benchmarking LLMs and agentic strategies on security challenges across HackTheBox, TryHackMe, PortSwigger Labs, Cybench,…

AI-native code security auditor on AgentField that proves exploitability with verdicts, traces, and actionable evidence.

Proof of concept for CVE-2024-24590

CAWODOG is a proof-of-concept project demonstrating how to protect Python-based AI models deployed on offline industrial machines. Across three…

Artefacts for blog post on finding CVE-2025-37899 with o3

A Proof-of-concept repository showing how an untrusted MCP server can steal literally everything...

Unified dashboard to monitor, govern, and audit AI agents in real-time. Enforce budgets, detect policy violations, and export compliance reports for…

Red Team AI Benchmark: Evaluating LLMs for authorized offensive-security tasks. Red Team AI Benchmark is a CLI model-evaluation benchmark. It…

Uses ChatGPT API, Bard API, and Llama2, Python-Nmap, DNS Recon, PCAP and JWT recon modules and uses the GPT3 model to create vulnerability reports…