
screenpipe
YC (S26) | Open Computer History | Record your screen continuously locally and provide context to your agents (Claude, Codex, Openclaw, Hermes,…

YC (S26) | Open Computer History | Record your screen continuously locally and provide context to your agents (Claude, Codex, Openclaw, Hermes,…

A full-stack AI Red Teaming platform securing AI ecosystems via Agent Scan, Skills Scan, MCP scan, AI Infra scan and LLM jailbreak evaluation.

A curated list of AI Security materials and resources for Pentesters, Bug Hunters, and Security Researchers.

Decentralized Autonomous Research Collective. Lean 4-verified, IPFS-backed, agent-native. Production paper: arXiv:2604.19792. Live at p2pclaw.com.

Cryptographically signed, replay-verifiable evidence layer for AI agents. Governs actions in the loop, produces Ed25519-signed receipts linked into a…

A serverless networking protocol designed for resilient state synchronization between autonomous agents in fragmented, low-bandwidth networks

Agent Control Protocol (ACP) — Official English specification. Cryptographically verifiable authorization architecture for autonomous AI agents.

ExploitGym is a large-scale, realistic benchmark built from real-world vulnerabilities designed to evaluate AI agents' ability to develop exploits.

Kernel-level eBPF sandbox for securing LLM agent tool calls made through the Model Context Protocol (MCP)


LLM-powered agent that autonomously fixes GitHub issues, finds cybersecurity vulnerabilities, and solves CTF challenges using configurable tool-use…

Automated Penetration Testing Agentic Framework Powered by Large Language Models

CVE-2025-67511: Tricking a Security AI Agent Into Pwning Itself

Research toolkit for analyzing AI agent behavioral patterns through multi-disciplinary corpus analysis. Parses session logs, runs 23 analytical…

Full-stack security OS for AI agents with five-layer defense-in-depth architecture covering foundation scan, input sanitization, cognition…

An intelligent agent for testing password strength, identifying vulnerabilities, and exploring patterns in password creation.

Turn any collection of documents into a knowledge graph. Extract entities and relationships via LLM, deduplicate with your approval. Map domains,…

Security scanner for MCP servers. Grades auth, permissions, injection risks, and tool safety. The Lighthouse of agent security.