Comparison

Galileo vs promptfoo

A like-for-like, spec-level comparison. Both entries are verified against their docs and repo.

Galileo

Guardrails and evaluation for production LLMs.

  • + Strong on hallucination and safety metrics
  • + Good for regulated industries
  • − Proprietary and enterprise-priced
  • − Less general-purpose tracing

promptfoo

Test and red-team LLM apps, prompts, and agents.

  • + Red-team scanning for prompt injection and jailbreaks
  • + Compares prompts and models side by side with assertion-based testing
  • − Test authoring requires upfront investment
  • − CLI-first; no hosted dashboard
Spec Galileo promptfoo
Role Evaluation Evaluation
Tags evaluation, guardrails, persona-data-scientist, persona-platform-engineer evaluation, guardrails, persona-data-scientist, persona-platform-engineer
License Proprietary MIT
Open source No Yes
Self-hostable No Yes
MCP support No No
Pricing paid free
Price Custom Free
Usage cost No model cost No model cost
Models multi multi
Languages python typescript
GitHub stars 22.9k
Last activity 2026-07-04