Skip to content
KitploitKITPLOIT
OutilsBlog
Soumettre
OutilsBlog
Soumettre

Outils de Hacking, PenTest et Cybersécurité pour votre Arsenal de Sécurité !

Kitploit est un répertoire d'outils de hacking, de cybersécurité et de pentesting. Découvrez les dernières mises à jour des projets pour trouver des vulnérabilités, analyser des systèmes, automatiser les tests et renforcer votre sécurité.

··Flux·Contact·Confidentialité·© 2026 Kitploit

Répertoire d'outils

Catégories

Voir toutes les catégories
Loading categories
Awesome-LLMs-for-Vulnerability-Detection — The community's most comprehensive, continuously-updated index of research on Large Language Models for software vulnerability detection — papers across function-level, repository-level, agentic, and smart-contract detection, plus datasets, benchmarks, and surveys. | Kitploit
Outils/GitHubGitHub/huhusmang/awesome-llms-for-vulnerability-detection
Static AnalysisVulnerability AnalysisCode AnalysisMachine LearningPapers & ResearchCurated ResourcesAI Security
GitHubhuhusmang/awesome-llms-for-vulnerability-detection

Awesome-LLMs-for-Vulnerability-Detection

The community's most comprehensive, continuously-updated index of research on Large Language Models for software vulnerability detection — papers across function-level, repository-level, agentic, and smart-contract detection, plus datasets, benchmarks, and surveys.

Voir le dépôt
1.2k108il y a 1 jourVérifié par Kitploit

Populaires

Voir tout →

Découvrez les outils les plus utilisés par notre communauté.

Explorer tous les outils

Parcourez notre collection d'outils

Voir tous les outils →
Partager
Contenu non disponible dans la langue demandée. Affichage de la version anglaise.

Awesome Large Language Models for Vulnerability Detection

A curated list of papers, projects, and agent skills on using LLMs for vulnerability detection and discovery.


📄 Papers

Only showing 2025 and later. For earlier work, see Papers Archive (2024 and earlier).

TitleVenueYearPaperGithub
VulnGym: Benchmarking Coding Agents for Repository-Level Vulnerability Detection2026linklink
VulTriage: Triple-Path Context Augmentation for LLM-Based Vulnerability Detection2026linklink
Synthesizing Multi-Agent Harnesses for Vulnerability Discovery2026linklink
QRS: A Rule-Synthesizing Neuro-Symbolic Triad for Autonomous Vulnerability Discovery2026link
Seclens: Role-specific Evaluation of LLM's for security vulnerablity detection2026linklink
Do Fine-Tuned LLMs Understand Vulnerabilities? An Investigation into the Semantic Trap2026link
Sifting the Noise: A Comparative Study of LLM Agents in Vulnerability False Positive FilteringISSTA2026link
AgenticSCR: An Autonomous Agentic Secure Code Review for Immature Vulnerabilities Detection2026link
LLM-based Vulnerability Detection at Project Scale: An Empirical Study2026link
MulVul: Retrieval-augmented Multi-Agent Code Vulnerability Detection via Cross-Model Prompt Evolution2026link
VulnLLM-R: Specialized Reasoning LLM with Agent Scaffold for Vulnerability Detection2025link
VULPO: Context-Aware Vulnerability Detection via On-Policy LLM Optimization2025link
VulInstruct: Teaching LLMs Root-Cause Reasoning for Vulnerability Detection via Security Specifications2025link
From Large to Mammoth: A Comparative Evaluation of Large Language Models in Vulnerability DetectionNDSS2025link
Benchmarking LLMs and LLM-based Agents in Practical Vulnerability Detection for Code RepositoriesACL2025linklink
A Systematic Literature Review on Detecting Software Vulnerabilities with Large Language Models2025linklink
LLMxCPG: Context-Aware Vulnerability Detection Through Code Property Graph-Guided Large Language ModelsUsenix2025linklink
CLeVeR: Multi-modal Contrastive Learning for Vulnerability Code RepresentationACL Findings2025linklink
Mono: Is Your "Clean" Vulnerability Dataset Really Solvable? Exposing and Trapping Undecidable Patches and Beyond2025linklink
Learning to Focus: Context Extraction for Efficient Code Vulnerability Detection with Language Models2025link
SV-TrustEval-C: Evaluating Structure and Semantic Reasoning in Large Language Models for Source Code Vulnerability AnalysisSP2025linklink
SecVulEval: Benchmarking LLMs for Real-World C/C++ Vulnerability Detection2025linklink
CVE-Bench: Benchmarking LLM-based Software Engineering Agent's Ability to Repair Real-World CVE VulnerabilitiesNAACL2025linklink
R2Vul: Learning to Reason about Software Vulnerabilities with Reinforcement Learning and Structured Reasoning Distillation2025linklink
Neuro-symbolic Static Analysis with LLM-generated Vulnerability Patterns2025link
Context-Enhanced Vulnerability Detection Based on Large Language Model2025link
Everything You Wanted to Know About LLM-based Vulnerability Detection But Were Afraid to Ask2025linklink
MOS: Towards Effective Smart Contract Vulnerability Detection through Mixture-of-Experts Tuning of Large Language Models2025link
Abundant Modalities Offer More Nutrients: Multi-Modal-Based Function-Level Vulnerability DetectionTOSEM2025linklink
Generative Large Language Model usage in Smart Contract Vulnerability Detection2025link
Closing the Gap: A User Study on the Real-world Usefulness of AI-powered Vulnerability Detection & Repair in the IDEICSE2025linklink
Vulnerability Detection with Code Language Models: How Far Are We?ICSE2025linklink
Combining Fine-Tuning and LLM-based Agents for Intuitive Smart Contract Auditing with JustificationsICSE2025link
LAMD: Context-driven Android Malware Detection and Classification with LLMs2025link
LLMs in Software Security: A Survey of Vulnerability Detection Techniques and Insights2025linklink
One-for-All Does Not Work! Enhancing Vulnerability Detection by Mixture-of-Experts (MoE)2025link

🚀 Projects

NameDescriptionGithub
OpenAnt (Knostic)LLM-powered multi-stage vulnerability discovery with adversarial verificationlink
DeepAuditMulti-agent AI red-team platform with Docker sandbox exploit validationlink
AutoCVEAutomated vulnerability detection and reporting with multi-agent architecturelink
DarkmoonOpen-source (GPL-3.0) autonomous AI pentest platform and MCP host; per-technology offensive sub-agents, Active Directory and Kubernetes coverage, 80+ orchestrated tools, evidence trail per findinglink
strixOpen-source autonomous AI penetration testing tool via piplink
deepsec (Vercel)Security harness for deep codebase vulnerability scanning with coding agentslink

🧩 Agent Skills

NameDescriptionLink
codex-security (OpenAI)Autonomous repo-level vulnerability scanning via Codex agentslink
defending-code-reference-harness (Anthropic)Reference skills for threat modeling, scanning, triage, and patching with Claude Codelink
security-audit-skill (Cloudflare)Six-phase security audit skill with parallel hunting agents and adversarial validationlink

📡 arxiv.md

Automated daily capture and update of Arxiv papers for specified keywords through workflows.

Acknowledgements

The project's Updated Arxiv Papers Daily workflow borrows from this project LLM4SE. I refactored its original code by using the arxiv library.

Télécharger l’outil