Skip to content
KitploitKITPLOIT
ToolsBlog
Submit
ToolsBlog
Submit

Hacking, PenTest, and Cybersecurity Tools for Your Security Arsenal!

Kitploit is a directory of hacking, cybersecurity, and pentesting tools. Discover the latest project updates to find vulnerabilities, analyze systems, automate testing, and strengthen your security.

··Feeds·Contact·Privacy·© 2026 Kitploit

Tool Directory

Categories

View all categories
Loading categories
Awesome-LLMs-for-Vulnerability-Detection — The community's most comprehensive, continuously-updated index of research on Large Language Models for software vulnerability detection — papers across function-level, repository-level, agentic, and smart-contract detection, plus datasets, benchmarks, and surveys. | Kitploit
Tools/GitHubGitHub/huhusmang/awesome-llms-for-vulnerability-detection
Static AnalysisVulnerability AnalysisCode AnalysisMachine LearningPapers & ResearchCurated ResourcesAI Security
GitHubhuhusmang/awesome-llms-for-vulnerability-detection

Awesome-LLMs-for-Vulnerability-Detection

Most Popular

View all →

Discover the most used tools by our community.

Explore all tools

Browse our collection of tools

View all tools →

The community's most comprehensive, continuously-updated index of research on Large Language Models for software vulnerability detection — papers across function-level, repository-level, agentic, and smart-contract detection, plus datasets, benchmarks, and surveys.

View Repository
1.2k1081 day agoReviewed by Kitploit
Share

Awesome Large Language Models for Vulnerability Detection

A curated list of papers, projects, and agent skills on using LLMs for vulnerability detection and discovery.


📄 Papers

Only showing 2025 and later. For earlier work, see Papers Archive (2024 and earlier).

TitleVenueYearPaperGithub
VulnGym: Benchmarking Coding Agents for Repository-Level Vulnerability Detection2026linklink
VulTriage: Triple-Path Context Augmentation for LLM-Based Vulnerability Detection2026linklink
Synthesizing Multi-Agent Harnesses for Vulnerability Discovery2026linklink
QRS: A Rule-Synthesizing Neuro-Symbolic Triad for Autonomous Vulnerability Discovery2026link
Seclens: Role-specific Evaluation of LLM's for security vulnerablity detection2026linklink
Do Fine-Tuned LLMs Understand Vulnerabilities? An Investigation into the Semantic Trap2026link
Sifting the Noise: A Comparative Study of LLM Agents in Vulnerability False Positive FilteringISSTA2026link
AgenticSCR: An Autonomous Agentic Secure Code Review for Immature Vulnerabilities Detection2026link
LLM-based Vulnerability Detection at Project Scale: An Empirical Study2026link
MulVul: Retrieval-augmented Multi-Agent Code Vulnerability Detection via Cross-Model Prompt Evolution2026link
VulnLLM-R: Specialized Reasoning LLM with Agent Scaffold for Vulnerability Detection2025link
VULPO: Context-Aware Vulnerability Detection via On-Policy LLM Optimization2025link
VulInstruct: Teaching LLMs Root-Cause Reasoning for Vulnerability Detection via Security Specifications2025link
From Large to Mammoth: A Comparative Evaluation of Large Language Models in Vulnerability DetectionNDSS2025link
Benchmarking LLMs and LLM-based Agents in Practical Vulnerability Detection for Code RepositoriesACL2025linklink
A Systematic Literature Review on Detecting Software Vulnerabilities with Large Language Models2025linklink
LLMxCPG: Context-Aware Vulnerability Detection Through Code Property Graph-Guided Large Language ModelsUsenix2025linklink
CLeVeR: Multi-modal Contrastive Learning for Vulnerability Code RepresentationACL Findings2025linklink
Mono: Is Your "Clean" Vulnerability Dataset Really Solvable? Exposing and Trapping Undecidable Patches and Beyond2025linklink
Learning to Focus: Context Extraction for Efficient Code Vulnerability Detection with Language Models2025link
SV-TrustEval-C: Evaluating Structure and Semantic Reasoning in Large Language Models for Source Code Vulnerability AnalysisSP2025linklink
SecVulEval: Benchmarking LLMs for Real-World C/C++ Vulnerability Detection2025linklink
CVE-Bench: Benchmarking LLM-based Software Engineering Agent's Ability to Repair Real-World CVE VulnerabilitiesNAACL2025linklink
R2Vul: Learning to Reason about Software Vulnerabilities with Reinforcement Learning and Structured Reasoning Distillation2025linklink
Neuro-symbolic Static Analysis with LLM-generated Vulnerability Patterns2025link
Context-Enhanced Vulnerability Detection Based on Large Language Model2025link
Everything You Wanted to Know About LLM-based Vulnerability Detection But Were Afraid to Ask2025linklink
MOS: Towards Effective Smart Contract Vulnerability Detection through Mixture-of-Experts Tuning of Large Language Models2025link
Abundant Modalities Offer More Nutrients: Multi-Modal-Based Function-Level Vulnerability DetectionTOSEM2025linklink
Generative Large Language Model usage in Smart Contract Vulnerability Detection2025link
Closing the Gap: A User Study on the Real-world Usefulness of AI-powered Vulnerability Detection & Repair in the IDEICSE2025linklink
Vulnerability Detection with Code Language Models: How Far Are We?ICSE2025linklink
Combining Fine-Tuning and LLM-based Agents for Intuitive Smart Contract Auditing with JustificationsICSE2025link
LAMD: Context-driven Android Malware Detection and Classification with LLMs2025link
LLMs in Software Security: A Survey of Vulnerability Detection Techniques and Insights2025linklink
One-for-All Does Not Work! Enhancing Vulnerability Detection by Mixture-of-Experts (MoE)2025link

🚀 Projects

NameDescriptionGithub
OpenAnt (Knostic)LLM-powered multi-stage vulnerability discovery with adversarial verificationlink
DeepAuditMulti-agent AI red-team platform with Docker sandbox exploit validationlink
AutoCVEAutomated vulnerability detection and reporting with multi-agent architecturelink
DarkmoonOpen-source (GPL-3.0) autonomous AI pentest platform and MCP host; per-technology offensive sub-agents, Active Directory and Kubernetes coverage, 80+ orchestrated tools, evidence trail per findinglink
strixOpen-source autonomous AI penetration testing tool via piplink
deepsec (Vercel)Security harness for deep codebase vulnerability scanning with coding agentslink

🧩 Agent Skills

NameDescriptionLink
codex-security (OpenAI)Autonomous repo-level vulnerability scanning via Codex agentslink
defending-code-reference-harness (Anthropic)Reference skills for threat modeling, scanning, triage, and patching with Claude Codelink
security-audit-skill (Cloudflare)Six-phase security audit skill with parallel hunting agents and adversarial validationlink

📡 arxiv.md

Automated daily capture and update of Arxiv papers for specified keywords through workflows.

Acknowledgements

The project's Updated Arxiv Papers Daily workflow borrows from this project LLM4SE. I refactored its original code by using the arxiv library.

Download Tool