promptfoo
by promptfoo
Tests and red-teams LLM apps, agents and RAG pipelines from declarative config, scanning for prompt injection, jailbreaks and other vulnerabilities in CI.
Skills
Declarative Eval Suites
Runs prompt and agent test cases from a config file, scoring outputs with assertions and model graders.
Automated Red Teaming
Generates adversarial attacks against an application to find prompt injection, jailbreaks and data leakage.
Model Comparison
Compares providers and prompt variants side by side on the same dataset to justify a model choice with data.
Related Agents
OpenSandbox
Runs AI-agent workloads in isolated Docker or Kubernetes sandboxes, exposing sandbox lifecycle, command, filesystem, an…
IDA Pro MCP
Bridges IDA Pro to LLM clients over MCP, exposing decompilation, disassembly, cross-references, type and stack edits, m…
Snyk
Developer security platform with AI-powered vulnerability detection, fix suggestions, and automated security testing ac…
SWE-agent
Takes a GitHub issue and autonomously fixes it with the language model of your choice. Also runs cybersecurity/CTF task…