What is TruLens?
An open-source library for systematically evaluating and tracing LLM-based applications, providing feedback functions and metrics to assess quality, safety, and relevance.
TruLens pricing
We don't have TruLens's full plan breakdown yet (its pricing page resisted automated reading). Here's what we could confirm. Always check live pricing for exact numbers.
What TruLens does
The capabilities that matter for ai evals testing, normalised so it lines up with every alternative. “-” means we haven't confirmed it, not that it's missing.
- Architecture model
- Open source CLI
- LLM as a judge prompt grading framework
- ✓
- Specialized rag metrics faithfulness context relevance
- ✓
- Deterministic regex and json schema assertions
- -
- Synthetic test dataset generation from documents
- -
- Ci cd github actions pipeline blocking gates
- -
- Red teaming and adversarial vulnerability scanning
- ✓
- Multi model side by side ab regression testing
- ✓
- Human in the loop hitl annotation UI
- -
- Dashboard analytics for metric drift over time
- ✓
- SOC2 type ii
- -
- Mit or apache permissive oss license
- -
- Pricing model
- Free open source
Platform & deployment
Independently observed- CLI
- Web
- Self-hosted
Integrations (2)
Independently observed- OpenAI
- GEPA
Security & compliance
Known vulnerabilities: 0 (0 in the last 12 months) sourcea count reflects scale & disclosure, not quality
TruLens alternatives
Other ai evals testing we track, ranked by the same independent score.
Compare TruLens
Side by side against other ai evals testing, attribute by attribute, with a source on every value.
The Vioscale score: one lens on the evidence
Not user reviews and not a paid placement: a confidence-weighted blend of the independent signals below (adoption, activity, security posture, and more), which you can sort and re-weight yourself. Vendors can correct their listing but can never move their rank, and stars are weighted low as a vanity metric. It is one way to read the evidence for TruLens, not the verdict.
| Signal | Score | Weight | Contribution | Evidence |
|---|---|---|---|---|
| Capabilities | 75 | 0.08 | 5.9 | ✓ |
| Development activity | 55 | 0.09 | 5.2 | ✓ |
| Release cadence | 92 | 0.05 | 4.8 | ✓ |
| Stars | 67 | 0.03 | 1.7 | ✓ |
| Integrations | 14 | 0.07 | 1.0 | ✓ |
| Dependent projects | 5 | 0.06 | 0.3 | ✓ |
| Security posture | 0 | 0.07 | 0.0 | - |
| Package downloads | 0 | 0.14 | 0.0 | - |
| Security score | 0 | 0.04 | 0.0 | - |
| Developer Q&A activity | 0 | 0.06 | 0.0 | - |
Computed . Re-weight it by intent, or see the full method.
All data & sourcesshow ↓
Every value we hold, with its source, retrieval date, and confidence. This is the evidence behind the score: don't trust it, verify it.
Activity
| Attribute | Value | Evidence |
|---|---|---|
| Commits last 30d | 55 | mediumsource · 2026-08-26 · 65% |
Adoption
Features
| Attribute | Value | Evidence |
|---|---|---|
| Capabilities | Pricing model: free_open_source · Architecture model: open_source_cli · Llm as a judge prompt grading framework: Yes · Dashboard analytics for metric drift over time: Yes · Multi model side by side ab regression testing: Yes · Red teaming and adversarial vulnerability scanning: Yes | mediumsource · 2026-08-21 · 60% |
Integrations
| Attribute | Value | Evidence |
|---|---|---|
| Count | 2 | mediumsource · 2026-08-21 · 60% |
Language
| Attribute | Value | Evidence |
|---|---|---|
| Primary | Python | highsource · 2026-08-26 · 90% |
License
| Attribute | Value | Evidence |
|---|---|---|
| Spdx | MIT | highsource · 2026-08-26 · 95% |
Pricing
Release
Reliability
| Attribute | Value | Evidence |
|---|---|---|
| Status page | Yes | mediumsource · 2026-08-21 · 60% |
Security
| Attribute | Value | Evidence |
|---|---|---|
| Vulnerabilities | Count: 0 · Source: https://advisories.ecosyste.ms/api/v1/advisories?ecosystem=pypi&package_name=trulens-eval&per_page=100 · Last 12m: 0 | highsource · 2026-08-26 · 90% |