Safety & Ethics — Free AI Tools

Bias detection, guardrails, explainability, evaluation.

Safety & Ethics

Bias detection, guardrails, explainability, evaluation.

8 tools
BIAS & GUARDRAILS4
IBM AI Fairness 360OSS
70+ fairness metrics + 10 mitigation algorithms
IBM Research
Free: Apache 2.0
NeMo GuardrailsOSS
Programmable LLM guardrails — PII, jailbreak, fact-check
NVIDIA
Free: Apache 2.0
DeepEvalOSS
30+ metrics + red teaming 40+ vulnerabilities
Confident AI
Free: Apache 2.0
Llama GuardOSS
Meta's open-source content safety classifier — 8B, runs locally
Meta
Free: Fully free open weights; runs locally
EXPLAINABILITY1
SHAPOSS
Gold standard feature attribution — 20K+ stars, industry standard
Community
Free: MIT
EVALUATION3
lm-eval-harnessOSS
60+ benchmarks — powers the Open LLM Leaderboard
EleutherAI
Free: MIT
RagasOSS
RAG evaluation — faithfulness, precision, recall metrics
Community
Free: Apache 2.0
PromptfooOSS
Test and evaluate LLM prompts — catch regressions before production
Promptfoo
Free: Fully free open-source; local-first
← Back to all tools