Skip to content
KitploitKITPLOIT
ツールブログ
提出
ツールブログ
提出

ハッキング、侵入テスト、サイバーセキュリティツールをあなたのセキュリティアーセナルに!

Kitploitはハッキング、サイバーセキュリティ、ペネトレーションテストのツールディレクトリです。最新のプロジェクトアップデートを見つけて、脆弱性の発見、システム分析、テストの自動化、セキュリティの強化を行いましょう。

··フィード·お問い合わせ·プライバシー·© 2026 Kitploit

ツールディレクトリ

カテゴリ

すべてのカテゴリを見る
Loading categories
Awesome-LLMs-for-Vulnerability-Detection — The community's most comprehensive, continuously-updated index of research on Large Language Models for software vulnerability detection — papers across function-level, repository-level, agentic, and smart-contract detection, plus datasets, benchmarks, and surveys. | Kitploit
ツール/GitHubGitHub/huhusmang/awesome-llms-for-vulnerability-detection
Static AnalysisVulnerability AnalysisCode AnalysisMachine LearningPapers & ResearchCurated ResourcesAI Security
GitHubhuhusmang/awesome-llms-for-vulnerability-detection

Awesome-LLMs-for-Vulnerability-Detection

人気

すべて見る →

コミュニティで最も使われているツールを見つけましょう。

すべてのツールを探索

ツールコレクションを閲覧

すべてのツールを見る →

The community's most comprehensive, continuously-updated index of research on Large Language Models for software vulnerability detection — papers across function-level, repository-level, agentic, and smart-contract detection, plus datasets, benchmarks, and surveys.

リポジトリを見る
1.2k1081日前Kitploit レビュー済み
共有
要求された言語のコンテンツは利用できません。英語版を表示しています。

Awesome Large Language Models for Vulnerability Detection

A curated list of papers, projects, and agent skills on using LLMs for vulnerability detection and discovery.


📄 Papers

Only showing 2025 and later. For earlier work, see Papers Archive (2024 and earlier).

TitleVenueYearPaperGithub
VulnGym: Benchmarking Coding Agents for Repository-Level Vulnerability Detection2026linklink
VulTriage: Triple-Path Context Augmentation for LLM-Based Vulnerability Detection2026linklink
Synthesizing Multi-Agent Harnesses for Vulnerability Discovery2026linklink
QRS: A Rule-Synthesizing Neuro-Symbolic Triad for Autonomous Vulnerability Discovery2026link
Seclens: Role-specific Evaluation of LLM's for security vulnerablity detection2026linklink
Do Fine-Tuned LLMs Understand Vulnerabilities? An Investigation into the Semantic Trap2026link
Sifting the Noise: A Comparative Study of LLM Agents in Vulnerability False Positive FilteringISSTA2026link
AgenticSCR: An Autonomous Agentic Secure Code Review for Immature Vulnerabilities Detection2026link
LLM-based Vulnerability Detection at Project Scale: An Empirical Study2026link
MulVul: Retrieval-augmented Multi-Agent Code Vulnerability Detection via Cross-Model Prompt Evolution2026link
VulnLLM-R: Specialized Reasoning LLM with Agent Scaffold for Vulnerability Detection2025link
VULPO: Context-Aware Vulnerability Detection via On-Policy LLM Optimization2025link
VulInstruct: Teaching LLMs Root-Cause Reasoning for Vulnerability Detection via Security Specifications2025link
From Large to Mammoth: A Comparative Evaluation of Large Language Models in Vulnerability DetectionNDSS2025link
Benchmarking LLMs and LLM-based Agents in Practical Vulnerability Detection for Code RepositoriesACL2025linklink
A Systematic Literature Review on Detecting Software Vulnerabilities with Large Language Models2025linklink
LLMxCPG: Context-Aware Vulnerability Detection Through Code Property Graph-Guided Large Language ModelsUsenix2025linklink
CLeVeR: Multi-modal Contrastive Learning for Vulnerability Code RepresentationACL Findings2025linklink
Mono: Is Your "Clean" Vulnerability Dataset Really Solvable? Exposing and Trapping Undecidable Patches and Beyond2025linklink
Learning to Focus: Context Extraction for Efficient Code Vulnerability Detection with Language Models2025link
SV-TrustEval-C: Evaluating Structure and Semantic Reasoning in Large Language Models for Source Code Vulnerability AnalysisSP2025linklink
SecVulEval: Benchmarking LLMs for Real-World C/C++ Vulnerability Detection2025linklink
CVE-Bench: Benchmarking LLM-based Software Engineering Agent's Ability to Repair Real-World CVE VulnerabilitiesNAACL2025linklink
R2Vul: Learning to Reason about Software Vulnerabilities with Reinforcement Learning and Structured Reasoning Distillation2025linklink
Neuro-symbolic Static Analysis with LLM-generated Vulnerability Patterns2025link
Context-Enhanced Vulnerability Detection Based on Large Language Model2025link
Everything You Wanted to Know About LLM-based Vulnerability Detection But Were Afraid to Ask2025linklink
MOS: Towards Effective Smart Contract Vulnerability Detection through Mixture-of-Experts Tuning of Large Language Models2025link
Abundant Modalities Offer More Nutrients: Multi-Modal-Based Function-Level Vulnerability DetectionTOSEM2025linklink
Generative Large Language Model usage in Smart Contract Vulnerability Detection2025link
Closing the Gap: A User Study on the Real-world Usefulness of AI-powered Vulnerability Detection & Repair in the IDEICSE2025linklink
Vulnerability Detection with Code Language Models: How Far Are We?ICSE2025linklink
Combining Fine-Tuning and LLM-based Agents for Intuitive Smart Contract Auditing with JustificationsICSE2025link
LAMD: Context-driven Android Malware Detection and Classification with LLMs2025link
LLMs in Software Security: A Survey of Vulnerability Detection Techniques and Insights2025linklink
One-for-All Does Not Work! Enhancing Vulnerability Detection by Mixture-of-Experts (MoE)2025link

🚀 Projects

NameDescriptionGithub
OpenAnt (Knostic)LLM-powered multi-stage vulnerability discovery with adversarial verificationlink
DeepAuditMulti-agent AI red-team platform with Docker sandbox exploit validationlink
AutoCVEAutomated vulnerability detection and reporting with multi-agent architecturelink
DarkmoonOpen-source (GPL-3.0) autonomous AI pentest platform and MCP host; per-technology offensive sub-agents, Active Directory and Kubernetes coverage, 80+ orchestrated tools, evidence trail per findinglink
strixOpen-source autonomous AI penetration testing tool via piplink
deepsec (Vercel)Security harness for deep codebase vulnerability scanning with coding agentslink

🧩 Agent Skills

NameDescriptionLink
codex-security (OpenAI)Autonomous repo-level vulnerability scanning via Codex agentslink
defending-code-reference-harness (Anthropic)Reference skills for threat modeling, scanning, triage, and patching with Claude Codelink
security-audit-skill (Cloudflare)Six-phase security audit skill with parallel hunting agents and adversarial validationlink

📡 arxiv.md

Automated daily capture and update of Arxiv papers for specified keywords through workflows.

Acknowledgements

The project's Updated Arxiv Papers Daily workflow borrows from this project LLM4SE. I refactored its original code by using the arxiv library.

ツールをダウンロード