Ad
 
Learn More

Open Source LLM Observability & Evaluation using Testing Library

A curated collection of the best open source tracing, monitoring, and evaluation platforms for LLM and AI-agent applications. using Testing Library.

Open source LLM engineering platform for AI-powered applications
  • Stars


    33,932
  • Last commit


    6 hours ago
  • License


    MIT
Langfuse provides tracing, evaluations, prompt management, and analytics to debug and improve LLM applications.

Open Source Alternative to:

Full observability for AI agents in production
  • Stars


    4,607
  • Last commit


    2 days ago
  • License


    MIT
Open-source platform for monitoring AI agents: captures traces, surfaces failure patterns, alerts on issues, and helps you verify fixes with automated evals.

Open Source Alternative to:

Simulation-based testing and evaluation for AI agents
  • Stars


    3,517
  • Last commit


    5 hours ago
  • License


    Apache-2.0
Tests AI agents through multi-turn simulations, LLM-based scoring, and production tracing so teams can ship reliable agents with confidence.

Open Source Alternative to:

Monitor, debug, and scale LLM applications with ease
  • Stars


    2,731
  • Last commit


    2 days ago
  • License


    Apache-2.0
Open-source observability platform for GenAI and LLM applications. Real-time monitoring, distributed tracing, prompt management, and AI model evaluation built on OpenTelemetry.

Open Source Alternative to:

LLM observability: cost, latency, and traces in one place
  • Stars


    12
  • Last commit


    7 days ago
  • License


    MIT
Drop-in observability platform for OpenAI, Anthropic, and Gemini that logs every request, tracks costs, traces agent workflows, and flags anomalies and PII.

Open Source Alternative to:

Favicon