
समुदाय का सबसे व्यापक, निरंतर अद्यतन होने वाला सूचकांक जो सॉफ्टवेयर भेद्यता पहचान के लिए लार्ज लैंग्वेज मॉडल्स पर शोध को दर्शाता है — फ़ंक्शन-स्तर, रिपॉजिटरी-स्तर, एजेंटिक, और स्मार्ट-कॉन्ट्रैक्ट पहचान से संबंधित पेपर, साथ ही डेटासेट, बेंचमार्क और सर्वेक्षण।
भेद्यता पहचान और खोज के लिए LLM उपयोग पर शोधपत्रों, परियोजनाओं और एजेंट कौशलों की एक चुनिंदा सूची।
केवल 2025 और उसके बाद के शोधपत्र दिखाए जा रहे हैं। पुराने कार्यों के लिए, देखें शोधपत्र संग्रह (2024 और पहले)।
| शीर्षक | प्रकाशन स्थल | वर्ष | शोधपत्र | Github |
|---|
| VulnGym: Benchmarking Coding Agents for Repository-Level Vulnerability Detection | 2026 | लिंक | लिंक | |
| VulTriage: Triple-Path Context Augmentation for LLM-Based Vulnerability Detection | 2026 | लिंक | लिंक | |
| Synthesizing Multi-Agent Harnesses for Vulnerability Discovery | 2026 | लिंक | लिंक | |
| QRS: A Rule-Synthesizing Neuro-Symbolic Triad for Autonomous Vulnerability Discovery | 2026 | लिंक | ||
| Seclens: Role-specific Evaluation of LLM's for security vulnerablity detection | 2026 | लिंक | लिंक | |
| Do Fine-Tuned LLMs Understand Vulnerabilities? An Investigation into the Semantic Trap | 2026 | लिंक | ||
| Sifting the Noise: A Comparative Study of LLM Agents in Vulnerability False Positive Filtering | ISSTA | 2026 | लिंक | |
| AgenticSCR: An Autonomous Agentic Secure Code Review for Immature Vulnerabilities Detection | 2026 | लिंक | ||
| LLM-based Vulnerability Detection at Project Scale: An Empirical Study | 2026 | लिंक | ||
| MulVul: Retrieval-augmented Multi-Agent Code Vulnerability Detection via Cross-Model Prompt Evolution | 2026 | लिंक | ||
| VulnLLM-R: Specialized Reasoning LLM with Agent Scaffold for Vulnerability Detection | 2025 | लिंक | ||
| VULPO: Context-Aware Vulnerability Detection via On-Policy LLM Optimization | 2025 | लिंक | ||
| VulInstruct: Teaching LLMs Root-Cause Reasoning for Vulnerability Detection via Security Specifications | 2025 | लिंक | ||
| From Large to Mammoth: A Comparative Evaluation of Large Language Models in Vulnerability Detection | NDSS | 2025 | लिंक | |
| Benchmarking LLMs and LLM-based Agents in Practical Vulnerability Detection for Code Repositories | ACL | 2025 | लिंक | लिंक |
| A Systematic Literature Review on Detecting Software Vulnerabilities with Large Language Models | 2025 | लिंक | लिंक | |
| LLMxCPG: Context-Aware Vulnerability Detection Through Code Property Graph-Guided Large Language Models | Usenix | 2025 | लिंक | लिंक |
| CLeVeR: Multi-modal Contrastive Learning for Vulnerability Code Representation | ACL Findings | 2025 | लिंक | लिंक |
| Mono: Is Your "Clean" Vulnerability Dataset Really Solvable? Exposing and Trapping Undecidable Patches and Beyond | 2025 | लिंक | लिंक | |
| Learning to Focus: Context Extraction for Efficient Code Vulnerability Detection with Language Models | 2025 | लिंक | ||
| SV-TrustEval-C: Evaluating Structure and Semantic Reasoning in Large Language Models for Source Code Vulnerability Analysis | SP | 2025 | लिंक | लिंक |
| SecVulEval: Benchmarking LLMs for Real-World C/C++ Vulnerability Detection | 2025 | लिंक | लिंक | |
| CVE-Bench: Benchmarking LLM-based Software Engineering Agent's Ability to Repair Real-World CVE Vulnerabilities | NAACL | 2025 | लिंक | लिंक |
| R2Vul: Learning to Reason about Software Vulnerabilities with Reinforcement Learning and Structured Reasoning Distillation | 2025 | लिंक | लिंक | |
| Neuro-symbolic Static Analysis with LLM-generated Vulnerability Patterns | 2025 | लिंक | ||
| Context-Enhanced Vulnerability Detection Based on Large Language Model | 2025 | लिंक | ||
| Everything You Wanted to Know About LLM-based Vulnerability Detection But Were Afraid to Ask | 2025 | लिंक | लिंक | |
| MOS: Towards Effective Smart Contract Vulnerability Detection through Mixture-of-Experts Tuning of Large Language Models | 2025 | लिंक | ||
| Abundant Modalities Offer More Nutrients: Multi-Modal-Based Function-Level Vulnerability Detection | TOSEM | 2025 | लिंक | लिंक |
| Generative Large Language Model usage in Smart Contract Vulnerability Detection | 2025 | लिंक | ||
| Closing the Gap: A User Study on the Real-world Usefulness of AI-powered Vulnerability Detection & Repair in the IDE | ICSE | 2025 | लिंक | लिंक |
| Vulnerability Detection with Code Language Models: How Far Are We? | ICSE | 2025 | लिंक | लिंक |
| Combining Fine-Tuning and LLM-based Agents for Intuitive Smart Contract Auditing with Justifications | ICSE | 2025 | लिंक | |
| LAMD: Context-driven Android Malware Detection and Classification with LLMs | 2025 | लिंक | ||
| LLMs in Software Security: A Survey of Vulnerability Detection Techniques and Insights | 2025 | लिंक | लिंक | |
| One-for-All Does Not Work! Enhancing Vulnerability Detection by Mixture-of-Experts (MoE) | 2025 | लिंक |
| नाम | विवरण | Github |
|---|---|---|
| OpenAnt (Knostic) | प्रतिकूल सत्यापन के साथ LLM-संचालित बहु-चरणीय भेद्यता खोज | लिंक |
| DeepAudit | Docker सैंडबॉक्स शोषण सत्यापन के साथ बहु-एजेंट AI रेड-टीम प्लेटफ़ॉर्म | लिंक |
| AutoCVE | बहु-एजेंट वास्तुकला के साथ स्वचालित भेद्यता पहचान और रिपोर्टिंग | लिंक |
| Darkmoon | ओपन-सोर्स (GPL-3.0) स्वायत्त AI पेंटेस्ट प्लेटफ़ॉर्म और MCP होस्ट; प्रति-प्रौद्योगिकी आक्रामक उप-एजेंट, Active Directory और Kubernetes कवरेज, 80+ ऑर्केस्ट्रेटेड उपकरण, प्रत्येक निष्कर्ष के लिए साक्ष्य निशान | लिंक |
| strix | pip के माध्यम से ओपन-सोर्स स्वायत्त AI पैठ परीक्षण उपकरण | लिंक |
| deepsec (Vercel) | कोडिंग एजेंटों के साथ गहन कोडबेस भेद्यता स्कैनिंग के लिए सुरक्षा हार्नेस | लिंक |
| नाम | विवरण | लिंक |
|---|---|---|
| codex-security (OpenAI) | Codex एजेंटों के माध्यम से स्वायत्त रिपॉजिटरी-स्तरीय भेद्यता स्कैनिंग | लिंक |
| defending-code-reference-harness (Anthropic) | Claude Code के साथ खतरा मॉडलिंग, स्कैनिंग, ट्राइएज और पैचिंग के लिए संदर्भ कौशल | लिंक |
| security-audit-skill (Cloudflare) | समानांतर शिकार एजेंटों और प्रतिकूल सत्यापन के साथ छह-चरणीय सुरक्षा ऑडिट कौशल | लिंक |
वर्कफ़्लो के माध्यम से निर्दिष्ट कीवर्ड के लिए Arxiv शोधपत्रों का स्वचालित दैनिक कैप्चर और अद्यतन।