
Awesome-LLMs-for-Vulnerability-Detection
The community's most comprehensive, continuously-updated index of research on Large Language Models for software vulnerability detection — papers…

The community's most comprehensive, continuously-updated index of research on Large Language Models for software vulnerability detection — papers…

Fully automatic censorship removal for language models

OGhidra bridges Large Language Models (LLMs) via Ollama with the Ghidra reverse engineering platform, enabling AI-driven binary analysis through…

Detects unknown jailbreak attacks in large vision-language models using hidden state analysis and autoencoders, with training and evaluation…

LLM-agent-powered concolic execution engine that instruments source code, summarizes path constraints in natural language, and generates test cases…

An alignment auditing agent capable of quickly exploring alignment hypothesis

Research code for extracting and training safety-awareness directions in multimodal LLMs to improve refusal behavior while limiting benign-task drift.

EmailXpose is an open source AI-powered email security system that detects phishing, spam, scams, malware, and social engineering attacks. It goes…

IDA plugin which queries language models to speed up reverse-engineering

🥂 Gracefully face hCaptcha challenge with multimodal large language model.

Automated security analysis pipeline that runs CodeQL queries on GitHub repositories and uses LLMs to classify and filter true vulnerabilities from…

A diagnostic framework for measuring LLM vulnerability to Affective Contextual Erosion (ACE) and related liminal attack vectors. **Delirium** is not…

Fully automatic censorship removal for language models

The code of VulTriage: Triple-Path Context Augmentation for LLM-Based Vulnerability Detection

AI / LLM Red Team Field Manual & Consultant’s Handbook

Code for ACL 2026 (main) paper "DeepGuard: Secure Code Generation via Multi-Layer Semantic Aggregation"

LLM powered fuzzing via OSS-Fuzz.

Evaluation framework that tests whether large language models follow invisible Unicode-encoded instructions embedded in normal-looking text, with…