🐢 Open-Source Evaluation & Testing library for LLM Agents
AI Red Team GitHub Repositories
Explore popular GitHub repositories tagged “ai-red-team”.
Compare stars, forks, and programming language using the same GitStar view as GitHub Trending.
Trending Repositories
The Python Risk Identification Tool for generative AI (PyRIT) is an open source framework built to empower security professionals and engineers to proactively identify risks in generative AI systems.
AI Red Teaming playground labs to run AI Red Teaming trainings including infrastructure.
Agentic LLM Vulnerability Scanner / AI red teaming kit 🧪
A powerful tool for automated LLM fuzzing. It is designed to help developers and security researchers identify and mitigate potential jailbreaks in their LLM APIs.
An offensive/defense security toolset for discovery, recon and ethical assessment of AI Agents
Autonomous AI pentesting engine, continuous offensive security across web, cloud, identity, CI/CD, IaC, databases, Active Directory, Kubernetes and IoT firmware. Agentic reasoning plus real exploit execution deliver proof-based vulnerabilities. Privacy gateway: the LLM never sees your real IPs, hosts or creds, nothing leaves your perimeter.
A comprehensive guide to adversarial testing and security evaluation of AI systems, helping organizations identify vulnerabilities before attackers exploit them.
LLM | Agentic | Security | Operations in one github repo with good links and pictures.
A professional AI security range for red teaming, vulnerability research, defensive validation, and hands-on AI/ML security training.
AI Security Platform: Defense (61 Rust engines + Micro-Model Swarm) + Offense (39K+ payloads)
AspGoat is an intentionally vulnerable ASP.NET Core application for learning and practicing web application security.
An open source plugin for enabeling claude to gain offensive pentesting capabilities
Open-Source autonomous security operations and red teaming agent built to help defenders investigate threats, analyze vulnerabilities, assess indicators of compromise, generate hardening guidance, and execute security research through an auditable agent workflow.
LMAP (large language model mapper) is like NMAP for LLM, is an LLM Vulnerability Scanner and Zero-day Vulnerability Fuzzer.
🤖🛡️🔍🔒🔑 Tiny package designed to support red teams and penetration testers in exploiting large language model AI solutions.
Toolkit for AI whitehats, internal red teams, llm bug bounty hunters and mlops - Adversarial testing for AI, LLMs, and Agents.
SOC-in-a-Box for AI purple teaming
Türkçe yapay zeka güvenliği için açık kaynak çatı: rehber serisi, uygulamalı akademi, araştırma & deneyler, 23+ araç ve 20 Hugging Face veri seti + model.
🛡️ Safe AI Agents through Action Classifier
AI/LLM Red Team Suite — Automated security testing toolkit for probing language models against prompt injection, jailbreaks, data extraction, and guardrail bypasses
90-day learning path from ML fundamentals to production AI security systems
Objective-driven adversarial testing framework for GenAI systems aligned with OWASP GenAI Top 10 risks.
Geometric AI governance and evaluation framework with a 14-layer security pipeline, semantic projection, and reproducible benchmark lanes.
This is my prompts for Lakera's Gandalf challenges
The SAPIEN Framework — an open standard (CC BY 4.0) for measuring AI behavioral safety (sycophantic drift), plus voigt-kampff, the FSL scoring CLI.
🤖 Test and secure AI systems with advanced techniques for Large Language Models, including jailbreaks and automated vulnerability scanners.