compare

DeepEval vs promptfoo

The same facts for both, read from GitHub every night, and the relation a person reviewed.

DeepEval vs promptfoo
FactDeepEvalpromptfoo
LanguagePythonTypeScript
LicenceApache-2.0MIT
Stars19k26k
Latestpython-v4.2.4code-scan-action-0.2.1
Last push2026-10-052026-10-06
Release cadenceabout 4 days between releasesabout 10 days between releases
Active contributors25+ commit authors on the default branch in the last 90 days51+ commit authors on the default branch in the last 90 days
Flagsnonenone
  • Both replace Braintrust. Alternatives to Braintrust →

    PartialDeepEvalPytest-style evaluations with LLM-as-judge metrics, in Python.

    PartialpromptfooEvals and red-teaming from a YAML config, run from the CLI or in CI, with a local web viewer.