Skip to content

Entry

OpenLLMetry

Appears in 7 awesome lists

OpenTelemetry-based instrumentation for LLM calls and agent steps: adds trace spans to every inference and tool call without modifying business logic. The cleanest way to bring the existing OTEL ecosystem (Grafana, Datadog, Jaeger) to a harness.

Open github.comtraceloop/openllmetry

Found in these lists

Awesome Harness Engineering

Section: Observability & Tracing · OpenTelemetry-based instrumentation for LLM calls and agent steps: adds trace spans to every inference and tool call without modifying business logic. The cleanest way to bring the existing OTEL ecosystem (Grafana, Datadog, Jaeger) to a harness.

FreshScore 88

Awesome LangChain

Section: Platforms · Open-source observability for your LLM application, based on OpenTelemetry

FreshScore 90

Awesome LLMOps

Section: Observability · OpenTelemetry-based observability and monitoring for LLM and agents workflows.

ActiveScore 75

Awesome local LLM

Section: Testing, Evaluation and Observability · an open-source observability for your LLM application, based on OpenTelemetry

FreshScore 87

Awesome Open Source AI

Section: 8. MLOps / LLMOps & Production · Open-source observability for GenAI/LLM applications based on OpenTelemetry with 25+ integration backends.

FreshScore 89

Awesome Production Machine Learning

Section: Evaluation and Monitoring · OpenLLMetry provides developers with deep visibility into Large Language Model applications through performance monitoring, execution tracing, and debugging capabilities.

FreshScore 92

awesome-python

Section: Monitoring and Observability · Open-source observability for your GenAI or LLM application, based on OpenTelemetry

FreshScore 81

Opik

Comet's open-source AI observability and evaluation platform: deep tracing of LLM calls, conversation logging, and agent activity, plus built-in eval metrics, prompt versioning, guardrails, and the Opik Agent Optimizer. Worth including because it unifies observability, verification, and…

In 16 listsDetails

Langfuse

LLM engineering platform for model tracing, prompt management, and application evaluation. Langfuse helps teams collaboratively debug, analyze, and iterate on their LLM applications such as chatbots or AI agents. (Demo, Source Code, Clients) MIT Docker

In 10 listsDetails

Phoenix

Open-source AI observability & evaluation platform (Arize) — OpenTelemetry-native tracing for agents, LLM-as-judge evals, versioned datasets & experiments for prompt regression testing, prompt management with version control and replay, plus an MCP endpoint so Claude Code/Cursor can query traces…

In 10 listsDetails

Langfuse

The most widely adopted self-hostable LLM observability platform: traces every agent step, manages prompt versions, and runs evals in one tool. Preferred over cloud-only alternatives when data residency or cost control is a constraint.

In 10 listsDetails

Deepchecks

Validation & testing of machine learning models and data during model development, deployment, and production. This includes checks and suites related to various types of issues, such as model performance, data integrity, distribution mismatches, and more.

In 8 listsDetails

Maxim AI

The agent simulation, evaluation, and observability platform helping product teams ship their AI applications with the quality and speed needed for real-world use.

In 8 listsDetails

Evidently

Interactive reports to analyze machine learning models during validation or production monitoring.

In 8 listsDetails

Helicone

Open-source LLM observability proxy (YC W23) with the largest open-source pricing database (300+ models). One-line proxy integration provides cost tracking, token monitoring, session tracing, and prompt versioning across providers. The AI Gateway component handles request routing and caching with…

In 8 listsDetails