Skip to content
KitploitKITPLOIT
工具博客
提交
工具博客
提交

黑客、渗透测试和网络安全工具,武装您的安全武器库!

Kitploit 是一个黑客、网络安全和渗透测试工具的目录。发现最新的项目更新,查找漏洞、分析系统、自动化测试并加强你的安全。

··订阅源·联系·隐私·© 2026 Kitploit

工具目录

分类

查看所有分类
Loading categories
Awesome-LLMs-for-Vulnerability-Detection — The community's most comprehensive, continuously-updated index of research on Large Language Models for software vulnerability detection — papers across function-level, repository-level, agentic, and smart-contract detection, plus datasets, benchmarks, and surveys. | Kitploit
工具/GitHubGitHub/huhusmang/awesome-llms-for-vulnerability-detection
Static AnalysisVulnerability AnalysisCode AnalysisMachine LearningPapers & ResearchCurated ResourcesAI Security
GitHubhuhusmang/awesome-llms-for-vulnerability-detection

Awesome-LLMs-for-Vulnerability-Detection

最受欢迎

查看全部 →

发现我们社区最常用的工具。

探索所有工具

浏览我们的工具集合

查看所有工具 →
分享

The community's most comprehensive, continuously-updated index of research on Large Language Models for software vulnerability detection — papers across function-level, repository-level, agentic, and smart-contract detection, plus datasets, benchmarks, and surveys.

查看仓库
1.2k1081天前Kitploit 审核通过
内容在请求的语言中不可用。显示英文版本。

Awesome Large Language Models for Vulnerability Detection

A curated list of papers, projects, and agent skills on using LLMs for vulnerability detection and discovery.


📄 Papers

Only showing 2025 and later. For earlier work, see Papers Archive (2024 and earlier).

TitleVenueYearPaperGithub
VulnGym: Benchmarking Coding Agents for Repository-Level Vulnerability Detection2026linklink
VulTriage: Triple-Path Context Augmentation for LLM-Based Vulnerability Detection2026linklink
Synthesizing Multi-Agent Harnesses for Vulnerability Discovery2026linklink
QRS: A Rule-Synthesizing Neuro-Symbolic Triad for Autonomous Vulnerability Discovery2026link
Seclens: Role-specific Evaluation of LLM's for security vulnerablity detection2026linklink
Do Fine-Tuned LLMs Understand Vulnerabilities? An Investigation into the Semantic Trap2026link
Sifting the Noise: A Comparative Study of LLM Agents in Vulnerability False Positive FilteringISSTA2026link
AgenticSCR: An Autonomous Agentic Secure Code Review for Immature Vulnerabilities Detection

🚀 Projects


🧩 Agent Skills


📡 arxiv.md

Automated daily capture and update of Arxiv papers for specified keywords through workflows.

Acknowledgements

The project's Updated Arxiv Papers Daily workflow borrows from this project LLM4SE. I refactored its original code by using the arxiv library.

下载工具
2026
link
LLM-based Vulnerability Detection at Project Scale: An Empirical Study2026link
MulVul: Retrieval-augmented Multi-Agent Code Vulnerability Detection via Cross-Model Prompt Evolution2026link
VulnLLM-R: Specialized Reasoning LLM with Agent Scaffold for Vulnerability Detection2025link
VULPO: Context-Aware Vulnerability Detection via On-Policy LLM Optimization2025link
VulInstruct: Teaching LLMs Root-Cause Reasoning for Vulnerability Detection via Security Specifications2025link
From Large to Mammoth: A Comparative Evaluation of Large Language Models in Vulnerability DetectionNDSS2025link
Benchmarking LLMs and LLM-based Agents in Practical Vulnerability Detection for Code RepositoriesACL2025linklink
A Systematic Literature Review on Detecting Software Vulnerabilities with Large Language Models2025linklink
LLMxCPG: Context-Aware Vulnerability Detection Through Code Property Graph-Guided Large Language ModelsUsenix2025linklink
CLeVeR: Multi-modal Contrastive Learning for Vulnerability Code RepresentationACL Findings2025linklink
Mono: Is Your "Clean" Vulnerability Dataset Really Solvable? Exposing and Trapping Undecidable Patches and Beyond2025linklink
Learning to Focus: Context Extraction for Efficient Code Vulnerability Detection with Language Models2025link
SV-TrustEval-C: Evaluating Structure and Semantic Reasoning in Large Language Models for Source Code Vulnerability AnalysisSP2025linklink
SecVulEval: Benchmarking LLMs for Real-World C/C++ Vulnerability Detection2025linklink
CVE-Bench: Benchmarking LLM-based Software Engineering Agent's Ability to Repair Real-World CVE VulnerabilitiesNAACL2025linklink
R2Vul: Learning to Reason about Software Vulnerabilities with Reinforcement Learning and Structured Reasoning Distillation2025linklink
Neuro-symbolic Static Analysis with LLM-generated Vulnerability Patterns2025link
Context-Enhanced Vulnerability Detection Based on Large Language Model2025link
Everything You Wanted to Know About LLM-based Vulnerability Detection But Were Afraid to Ask2025linklink
MOS: Towards Effective Smart Contract Vulnerability Detection through Mixture-of-Experts Tuning of Large Language Models2025link
Abundant Modalities Offer More Nutrients: Multi-Modal-Based Function-Level Vulnerability DetectionTOSEM2025linklink
Generative Large Language Model usage in Smart Contract Vulnerability Detection2025link
Closing the Gap: A User Study on the Real-world Usefulness of AI-powered Vulnerability Detection & Repair in the IDEICSE2025linklink
Vulnerability Detection with Code Language Models: How Far Are We?ICSE2025linklink
Combining Fine-Tuning and LLM-based Agents for Intuitive Smart Contract Auditing with JustificationsICSE2025link
LAMD: Context-driven Android Malware Detection and Classification with LLMs2025link
LLMs in Software Security: A Survey of Vulnerability Detection Techniques and Insights2025linklink
One-for-All Does Not Work! Enhancing Vulnerability Detection by Mixture-of-Experts (MoE)2025link
NameDescriptionGithub
OpenAnt (Knostic)LLM-powered multi-stage vulnerability discovery with adversarial verificationlink
DeepAuditMulti-agent AI red-team platform with Docker sandbox exploit validationlink
AutoCVEAutomated vulnerability detection and reporting with multi-agent architecturelink
DarkmoonOpen-source (GPL-3.0) autonomous AI pentest platform and MCP host; per-technology offensive sub-agents, Active Directory and Kubernetes coverage, 80+ orchestrated tools, evidence trail per findinglink
strixOpen-source autonomous AI penetration testing tool via piplink
deepsec (Vercel)Security harness for deep codebase vulnerability scanning with coding agentslink
NameDescriptionLink
codex-security (OpenAI)Autonomous repo-level vulnerability scanning via Codex agentslink
defending-code-reference-harness (Anthropic)Reference skills for threat modeling, scanning, triage, and patching with Claude Codelink
security-audit-skill (Cloudflare)Six-phase security audit skill with parallel hunting agents and adversarial validationlink