Promptfoo Evals
Teaches AI coding agents to create and maintain promptfoo eval suites. Encodes best practices like deterministic assertions first, file-based test organization, correct environment variable syntax, and proper hallucination/faithfulness checks. Includes a cheatsheet covering 30+ assertion types, provider patterns for OpenAI/Anthropic/Google/Bedrock/HTTP/Python/JS, and test patterns for JSON validation, dataset-driven scaling, and CI-friendly runs.
Works with: Claude Code, Cursor, Claude Desktop, Codex CLI, Gemini CLI, Cline, Windsurf, VS Code
Category: Dev Tools & CI — see all ranked ›
Work: Model evaluation
Who it is for: AI engineer
- tashan score: 81.0
- Adoption: 1 repos
- Health: active
- GitHub stars: 23,659
- Contributors: 30
- License: MIT
Security audit
Not scanned yet. We audit npm-published capabilities for known advisories, install-time scripts and permission surface; this one has no npm package we can resolve, or has not reached the queue.
source ↗ · plugin:promptfoo/promptfoo/promptfoo-evals
Already running this? npx tashan-cli doctor checks your whole config against the Index — how it works ›