Observability & tracing
tracing, logging and debugging agents
raga-ai-hub/RagaAI-Catalyst raga-ai-hub/RagaAI-Catalyst is the clear leader; if you want a single self‑hosted platform for multi‑agent tracing, timeline analytics, evaluation and guardrails, it’s the most mature choice. Pick other projects only when you need lighter instrumentation, OTel-native traces, production-to‑CI replay tests, per‑line provenance, or fully local developer UIs.
Which one matches your setup?
Answer any of the questions — the shortlist updates as you go. Recommendations come from the capability passports below, nothing else.
Comparison matrix
Axes are extracted from each project's docs by our review pipeline; the maturity score is computed from stars, growth and commit activity — not an opinion. Click a column to sort.
| Runs in | Models | Context | Cost to run | |||||
|---|---|---|---|---|---|---|---|---|
⭐ 2.7k +19/7d | Web appCLICoding-agent plugin | BYOKOllama / local | Full-stack traces | Rule engine filters | ●●●●● | Self-hostable | ●●●●● | Self-hosted; you pay for LLM/provider API usage. |
⭐ 5.8k +6/7d | CLIWeb app | BYOKOpenAIAnthropicFixed provider | Runtime tracing | None mentioned | ●●●●● | Self-hostable | ●●●●● | AgentOps API key for hosted dashboard; can self-host |
⭐ 1.2k | GitHub ActionCICLIWeb app | BYOKOpenAIAnthropic | No repo context | Clustering & gating | ●●●●● | Self-hostable / local | ●●●●● | Self-hosted; CI replays use recorded fixtures (no model spend). |
⭐ 1.2k | GitHub ActionCICLIWeb app | BYOKOpenAIAnthropic | Per-run only | Clustering & gating | ●●●●● | Self-hostable | ●●●●● | CI replay $0; live judges use your model API key. |
⭐ 639 | CLIWeb app | AnthropicOpenAIFixed provider | Whole-repo analysis | Minimal signals | ●●●●● | Fully local | ●●●●● | Free, local |
⭐ 468 +8/7d | CLI | OpenAIAnthropicFixed provider | Session events only | None mentioned | ●●●●● | Fully local | ●●●●● | Free, local |
⭐ 16.2k +11/7d | CLIWeb appCI | BYOK | Diff/file only | Configurable thresholds | ●●●●● | Self-hostable | ●●●●● | Requires RagaAI account; external LLM usage billed to your provider. |
⭐ 789 +3/7d | CLIIDECoding-agent plugin | AnthropicOpenAIFixed provider | Workspace snapshot | Ignore & dedupe | ●●●●● | Self-hostable | ●●●●● | Free local CLI; external model/API usage may incur costs |
Capability profiles
Six axes, 0–5 each. The shape tells you the strategy: a wide hexagon is a generalist, a spike is a specialist.
OpenLIT provides an OpenTelemetry-native, open-source AI observability platform combining traces, built-in evaluations, a rule engine, prompt management and guardrails in one stack.
Provides end-to-end agent session replays, LLM cost tracking and debugging across many agent frameworks with minimal instrumentation.
Promotes real failing production traces into hermetic, replayable regression tests that run offline in CI and block PRs.
Promotes real failing production traces into hermetic, replayable regression cases that run offline in CI and block PRs, instead of relying on hand-authored datasets.
Provides a local, real-time observability map of coding agents' plans, tool calls, and file edits without any cloud service, account, or telemetry.
Visualizes multiple coding agents as a whimsical, glanceable pixel-art terminal 'office' with animated coworkers and per-agent session telemetry.
Combines agent/LLM tracing, evaluation, guardrails and red‑teaming with a self‑hosted dashboard and execution-timeline analytics for multi-agent debugging.
Provides content-addressed, per-step audit trails and per-line 'blame' tied to the exact prompt and conversation that produced a change—something traditional VCS and most tools do not track.
All repositories (8)
Python SDK for Agent AI Observability, Monitoring and Evaluation Framework. Includes features like agent, llm…
Python SDK for AI agent monitoring, LLM cost tracking, benchmarking, and more. Integrates with most LLMs and a…
Open source platform for AI Engineering: OpenTelemetry-native LLM Observability, GPU Monitoring, Guardrails, E…
Trace-native CI/CD for AI agents — production failures become regression tests that block the PR. Auto-detect,…
Trace-native CI/CD for AI agents — production failures become regression tests that block the PR. Auto-detect,…
Version control for AI agents — track what your agent did, blame any line to a prompt, inspect any step.
An infinite canvas for your AI coding agents. Every repo is a region on one zoomable live map - watch Claude C…