Evidently vs OpenLLMetry
An independent, evidence-based trust comparison of Evidently and OpenLLMetry, two Observability & Evaluation projects in the HVTracker registry. Scores come from public, checkable signals — supply-chain provenance, OSSF Scorecard, maintenance, and adoption — not popularity.
OpenLLMetry leads on trust — 83.8/100 (Grade A) vs 76.3/100 (Grade B), a 7.5-point gap. Note: OpenLLMetry's evidence coverage is thinner (coverage B vs A) — its score rests on fewer independent signal types. Full breakdown below.
| Signal | Evidentlyevidentlyai/evidently | OpenLLMetrytraceloop/openllmetry |
|---|---|---|
| HVTrust score | 76.3 | 83.8 |
| Evidence grade | B | A |
| Coverage grade | A | B |
| Overall rank | #71 | #48 |
| Rank in Observability & Evaluation | #4 | #3 |
| GitHub stars | 7.8k | 7.4k |
| Last updated | today | 1d ago |
| Build provenance | Yes | Yes |
| OSSF Scorecard | 3.3 / 10 | 6.7 / 10 |
| License | Apache-2.0 | Apache-2.0 |
| Downloads | 296k/wk | 566k/wk |
| Trust dimensions (points earned) | ||
| Safety / integrity / 25 | 16.1 | 19.9 |
| Identity & provenance / 18 | 18.0 | 18.0 |
| Transparency / 17 | 11.3 | 14.2 |
| Maintenance / 20 | 12.0 | 11.9 |
| Adoption / 20 | 16.6 | 17.0 |
| Runtime capability surface (full matrix) | ||
| MCP server | — | Implemented |
| External providers | 2 — OpenAI, Postgres | 5 — Amazon Bedrock, Anthropic, Google Gemini, … |
| Requires API keys | No | No |
| Plugin surface | — | extensions |
| Provenance drift | Match | Match |
Open in the live compare tool →
Evidently profile
OpenLLMetry profile
More Observability & Evaluation →
How to read this: HVTrust (0–100) weighs supply-chain signals (provenance, OSSF Scorecard, signed commits, open license) alongside real-world adoption, scaled by an evidence-confidence factor. Grade bands: A ≥ 80, B ≥ 65, C ≥ 50, D < 50. Signals refresh daily. Full methodology v4.2 →