
challenge_repo
Structured challenge repository offering bounties for breaking a neural-network-based random number generator validated against NIST SP 800-22 and SP…

Structured challenge repository offering bounties for breaking a neural-network-based random number generator validated against NIST SP 800-22 and SP…
Multi-agent automated context management for long horizon tasks in local AI Agents

Hybrid machine-learning pipelines for detecting SQL injection in web traffic, combining DistilBERT and BERT-GNN models with adversarial training and…

Simulates CVE-2024-38063 TCP/IP remote code execution attack, captures network traffic with TShark, and trains a machine learning model to detect…

Implements machine and deep learning methods for indoor UWB jammer localization, including hyperparameter optimization, classification, and…

An automated, high-precision zero-shot evaluation pipeline for OpenAI's CLIP model on CIFAR-10. Features 88.80% accuracy, Safetensors security…

Open framework for RL-based prompt injection red teaming, with a shared trainer, curriculum learning, and benchmarks like AgentDojo, InjecAgent, and…

Benchmark harness measuring where prompt injection defenses fire in tool-using LLM agent pipelines, tracking canary tokens across exposed, persisted,…

A lifecycle benchmark for black-box LLM extraction attacks, defenses, and adaptive attacks.

Re-evaluating prompt-injection detectors on LLM agent tool outputs (paper draft, scripts, scores)