
OMLASP
Framework for auditing machine learning algorithms against adversarial attacks and biases, providing educational tools and an academic paper to…

Framework for auditing machine learning algorithms against adversarial attacks and biases, providing educational tools and an academic paper to…

Fully automatic censorship removal for language models

Open-source email filtering framework that detects spam and phishing using content analysis, header checks, Bayesian scoring, and DNS blocklists.

Open source entropy based invalid traffic detection and pre-bid filtering.

A Python pickling decompiler and static analyzer

A machine learning tool that ranks strings based on their relevance for malware analysis.

Research pipeline for detecting latent indirect prompt-injection exposure signals in agentic LLMs via hidden-state probing, including trace…

Streaming machine learning library for incremental learning on data streams, providing online estimators, drift and anomaly detection, pipelines,…

Re-play Security Events

An Open-Source Package for Textual Adversarial Attack.

Linux system-call monitor using ptrace to trace file, process, network, and memory activity, with namespace isolation and machine learning…

Research implementation for mitigating adaptive prompt injections via on-policy distillation, with training recipes and evaluators for SEP, PISmith,…

Collection of CVE(work) on tenserflow binary pwning it

ML-driven threat detection and continuous monitoring platform built for federal zero trust architectures.

DGA Domain Detection using Bigram Frequency Analysis

Curated list of backdoor learning papers, surveys, and toolboxes, organizing poisoning-based attacks and defenses in deep learning for researchers…

Programmable guardrails for LLM chat apps: enforce input/output rails, block jailbreaks and prompt injections, detect hallucination, and mask…

Research code for detecting and detoxifying backdoors in text-to-image diffusion models, with pipelines for Stable Diffusion v1.4, v1.5, XL, and 3…