Latitude vs Opik
Both are Grade B and 1.6 points apart, so choose on what you weigh most. Opik leads on adoption and rests on broader evidence.
Open-source observability for AI agents. Find where your agents fail, dispatch your coding agent to fix it, and verify the fix against real traces.
Latitude doesn't lead on any scored dimension in this pair.
Debug, evaluate, and monitor your LLM applications, RAG systems, and agentic workflows with comprehensive tracing, automated evaluations, and production-ready dashboards.
Choose Opik if adoption matters most.
- +4.2Adoption: 441.1k weekly downloads against 5.4k
- A vs BEvidence coverage: 4 of 5 independent signal types, against 3
Where they differ
3 dimensions identical: Identity 18.0 · Transparency 8.5 · Maintenance 19.9 · Full evidence table
An independent, evidence-based trust comparison of Latitude and Opik, two Observability & Evaluation projects in the HVTracker registry. Scores come from public, checkable signals — supply-chain provenance, OSSF Scorecard, maintenance, and adoption — not popularity.
Full evidence
| Signal | Latitudelatitude-dev/latitude-llm | Opikcomet-ml/opik |
|---|---|---|
| HVTrust score | 76.9 | 78.5 |
| Evidence grade | B | B |
| Coverage grade | B | A |
| Overall rank | #210 | #185 |
| Rank in Observability & Evaluation | #8 | #7 |
| GitHub stars | 4.7k | 22.5k |
| Last updated | 2d ago | 1d ago |
| Build provenance | Yes | Yes |
| OSSF Scorecard | — | — |
| License | MIT | Apache-2.0 |
| Downloads | 5k/wk | 441k/wk |
| Trust dimensions (points earned) | ||
| Safety / integrity / 25 | 12.0 | 12.2 |
| Identity & provenance / 18 | 18.0 | 18.0 |
| Transparency / 17 | 8.5 | 8.5 |
| Maintenance / 20 | 19.9 | 19.9 |
| Adoption / 20 | 13.8 | 18.0 |
| Runtime capability surface (full matrix) | ||
| MCP server | Implemented | Implemented |
| External providers | — | 4 — Amazon Bedrock, Anthropic, Multi-provider (LiteLLM), … |
| Requires API keys | Yes | No |
| Plugin surface | plugins | extensions |
| Provenance drift | Match | Partial |
How to read this: HVTrust (0–100) weighs supply-chain signals (provenance, OSSF Scorecard, signed commits, open license) alongside real-world adoption, scaled by an evidence-confidence factor. Grade bands: A ≥ 80, B ≥ 65, C ≥ 50, D < 50. Signals refresh daily. Full methodology v4.4 →