These tools integrates with

DeepTeamvsDeepEval

OSS red-teaming framework for LLMs and agents versus LLM evaluation framework — 14+ metrics

Compare interactively in Explore →

Choose DeepTeam when…

  • You need to red-team an LLM or agent against a broad, named vulnerability taxonomy
  • You want OWASP LLM Top 10 / NIST AI RMF mapped attack coverage out of the box
  • You're already using DeepEval and want a matching adversarial-testing tool

Choose DeepEval when…

  • You want a pytest-style framework for LLM testing
  • Unit-test-like evals for LLM outputs fit your workflow
  • You need RAG-specific metrics like faithfulness and relevancy

Side-by-side comparison

Field
DeepTeam
DeepEval
Category
Prompt & Eval
Prompt & Eval
Type
Open Source
Open Source
Free Tier
✓ Yes
✓ Yes
Pricing Plans
OSS: Free
GitHub Stars
2,300
5,500
Health
85 Active
95 Active

DeepTeam

Adversarial red-teaming framework covering 50+ vulnerability types and 20+ attack methods, mapped to the OWASP LLM Top 10 and NIST AI RMF. Built by Confident AI, the team behind DeepEval, as a dedicated security-testing counterpart rather than a general eval library.

DeepEval

Open-source evaluation framework with 14+ metrics including faithfulness, relevancy, and hallucination detection. Integrates with CI/CD.

Shared Connections1 tool both integrate with

Only DeepTeam (1)

DeepEval

Only DeepEval (7)

LangfuseRAGASPromptFooOpenAI APITruLensGalileoDeepTeam

Explore the full AI landscape

See how DeepTeam and DeepEval fit into the bigger picture — 256 tools, 553 relationships, all mapped.

Open in Explore →