LangWatch
FreeAn Open Source tool for observing, evaluating and optimising your llm apps and prompts, which supports LangChain out of the box! 
About LangWatch
LangWatch is an open-source platform for LLM evaluations and AI agent testing. It enables teams to test, simulate, evaluate, and monitor LLM-powered agents end-to-end—before release and in production. It offers end-to-end agent simulations with realistic scenarios, a unified loop for tracing, dataset creation, evaluation, and prompt optimization, and is built on open standards like OpenTelemetry (OTLP) for framework and provider agnosticism. It includes an AI Gateway with a proxy for governance, cost control, and fallback across providers. Collaboration features include run reviews, annotation queues, and GitHub integration for versioning prompts. LangWatch supports cloud and self-hosted deployments.
Key Features
Pros & Cons
- Open-source and self-hostable, giving full data control
- No vendor lock-in due to open standards (OTLP)
- Provides end-to-end visibility from simulation to production
- Integrated AI gateway for cost management and fallback
- Designed for team collaboration with annotation and Git integration