You signed in with another tab or window. Reload to refresh your session.You signed out in another tab or window. Reload to refresh your session.You switched accounts on another tab or window. Reload to refresh your session.Dismiss alert
Research repository containing safe evaluation prompts for LLM alignment testing. Study how to jailbreak claude (including opus jailbreak & claude 5.0 jailbreak methods), analyze claude code jailbreak boundaries, and understand system instructions.
LLM Attack Testing Toolkit is a structured methodology and mindset framework for testing Large Language Model (LLM) applications against logic abuse, prompt injection, jailbreaks, and workflow manipulation.
LLM Sentinel Red Teaming Platform is an enterprise-grade framework for automated security testing of Large Language Models, detecting vulnerabilities such as jailbreaks, prompt injection, and system prompt leakage across multiple providers, with structured attack orchestration, risk scoring, and security reporting to harden models before production
Advanced Claude Prompt engineering repository focused on LLM safety research, system prompt analysis, and creative boundary exploration. Multi-version compatible guidelines for advanced AI behavior understanding.
AI testing portfolio project - LLM output testing, prompt injection, and non-deterministic testing patterns using Python, pytest, DeepEval, and Claude API
Promptfoo is an open-source CLI and TypeScript/Node.js library for evaluating, red-teaming, and security-testing LLM applications, agents, and RAG pipelines. It runs deterministic prompt evals with model-graded and rule-based assertions, generates dynamic adversarial attack probes across 50+ vulnerability categories (prompt injection, jailbreaks…
Mindgard — independent third-party profile of a public API surface, by API Evangelist. Mindgard is a UK-based offensive AI security company (London/Lancaster, spun out of Lancaster University) that provides an automated AI red-teaming and security testing platform for large language models, AI agents, and generative AI systems.