promptfoo
Official@promptfoo · United States of America
Test your LLM apps
Agent Skills by promptfoo
Showing 7 vetted skills indexed across 1 GitHub repositories.
promptfoo-evals
Write, run, and QA promptfoo eval suites with assertions and model-graded rubrics.
search-params
Manages URL search param and hash state with correct browser history behavior in React.
token-skill
Returns a fixed response token when explicitly invoked by the user.
standards-check
Checks that a project contains required standard files like README.md.
code-review
Reviews code for bugs, security vulnerabilities, and best practice violations.
discount-review
Inspects a discount policy fixture using a Python helper script and review checklist.
redteam-plugin-development
Standardize redteam plugin development and grading with rubric structures.
Frequently Asked Questions About promptfoo
FAQPage SchemaWhat specific security tasks does this framework enable?▼
This framework enables the creation of custom redteam plugins and the implementation of structured grading rubrics. It allows engineers to systematically evaluate model outputs against defined security benchmarks, ensuring consistent detection of vulnerabilities and adversarial patterns during the testing phase.
Which technical personas benefit from these redteaming capabilities?▼
Security engineers, model evaluators, and compliance officers benefit from these capabilities. The framework is designed for technical teams responsible for validating model safety, establishing rigorous testing standards, and maintaining consistent security posture across production-grade deployments.
What are the primary prerequisites for implementing these security rubrics?▼
Implementation requires a defined set of security test cases and established criteria for model output evaluation. Users must define their specific threat models and rubric structures to align with organizational safety requirements before deploying the plugin development environment.