AIAI Governance StackFree kit

Promptfoo vs Giskard

Both compete in Red-Teaming & AI Security. Promptfoo positions itself as “Open-source tool for evaluating and red-teaming LLM apps, agents, and RAG systems”, while Giskardleads with “Open-source and enterprise platform for testing and red-teaming LLM agents”. The table below compares what each publishes.

Where Promptfoo pulls ahead

Developer teams wanting free, open-source, CI/CD-integrated evaluation and red teaming of LLM apps

Where Giskard pulls ahead

Publishes support for EU AI Act, NIST AI RMF, which Promptfoo does not. ML, quality, and risk teams wanting open-source-first, framework-aligned LLM testing and continuous red teaming

PositioningOpen-source tool for evaluating and red-teaming LLM apps, agents, and RAG systemsOpen-source and enterprise platform for testing and red-teaming LLM agents
CategoryRed-Teaming & AI SecurityRed-Teaming & AI Security
FrameworksNone publishedEU AI Act, NIST AI RMF
DeploymentOpen-source, SaaS, APIOpen-source, SaaS, On-prem, API
Built forData Science / ML, SecurityData Science / ML, Risk, Compliance
Founded20232021
HeadquartersUSAParis, France
OwnershipAcquired by OpenAI (2026); previously VC-backedIndependent, venture-backed (Y Combinator alumnus)
Funding~$23.4M raised prior to acquisition; $5M seed (a16z, 2024) and $18.4M Series A led by Insight Partners (2025)Seed funding (reported ~$2-3M+); investors include Y Combinator, Elaia, and others
PricingOpen-source (MIT license), free; paid enterprise platformOpen-source library (free); Giskard Hub commercial enterprise subscription
Key capabilities
  • LLM evaluation and benchmarking
  • Automated red teaming and vulnerability scanning
  • Declarative test configs
  • Prompt and model comparison
  • CI/CD integration
  • Local/self-hosted execution
  • Open-source LLM/model vulnerability scanning
  • Automated test-suite generation
  • Continuous red teaming
  • Hallucination and prompt-injection testing
  • Robustness and bias evaluation
  • Business-domain test management (Giskard Hub)
IntegrationsOpenAI, Anthropic, Google Gemini, DeepSeek, CI/CD pipelines, GitHubHugging Face, LangChain, MLflow, Common ML frameworks, Major LLM providers via API
Notable customersOpenAI, Anthropic, Fortune 500 enterprisesNone published
Best forDeveloper teams wanting free, open-source, CI/CD-integrated evaluation and red teaming of LLM appsML, quality, and risk teams wanting open-source-first, framework-aligned LLM testing and continuous red teaming
LimitationsDeveloper-oriented and requires engineering effort to configure; broader governance/compliance features live in the paid enterprise tier.Testing/evaluation focus rather than inline runtime enforcement; smaller company and funding base; enterprise features concentrated in the paid Hub tier

Which should you shortlist?

Choose Promptfoo if developer teams wanting free, open-source, CI/CD-integrated evaluation and red teaming of LLM apps

Choose Giskard if mL, quality, and risk teams wanting open-source-first, framework-aligned LLM testing and continuous red teaming

Neither is a substitute for a governance program. Whichever you pick, you still need people who can define the policies the tool enforces.

AI Governance Tool Selection Kit

A vendor-comparison worksheet plus EU AI Act, NIST AI RMF and ISO/IEC 42001 requirement checklists — so you can shortlist tools against the obligations that actually apply to you.

Free. No spam — unsubscribe anytime.