Skip to main content
Understand what your AI agents are doing, measure their performance, and improve them in production. With OpenLIT teams can Trace LLM and agent calls, compare models, iterate on prompts, detect regressions, and optimize cost and performance using real production data.
Prompt management, prompt versioning, and model testing tools for teams building AI applications:
  • Prompt versioning and deployment - Prompt Hub treats prompts as versioned artifacts. Edit, version, and deploy to any environment without code changes. Roll back instantly when needed.
  • Collaborate on prompts - Edit prompts as a team with full change history. Every version is tracked and linked to the evaluations and traces it produced, giving you full traceability from prompt to output.
  • Model comparison and A/B testing - OpenGround runs side-by-side LLM prompt testing across models, comparing cost, latency, and output quality so you pick the right model before you ship.

Get Started

Instrument an AI Agent

Production-ready AI Observability in 2 steps with zero code changes

Observe coding agents

Track spend and usage for Claude Code, Cursor, & more across teams

Deploy OpenLIT

Self-host the full platform with Docker Compose or Helm

Evaluate LLM responses

Score live traces automatically with prebuilt LLM-as-a-judge evaluators

Frequently asked questions

OpenLIT is an open-source AI Engineering platform. It traces LLM and agent calls via OpenTelemetry, tracks cost and token usage, runs automated evaluations, and manages prompts - all self-hostable.
Yes. OpenLIT is fully open source and self-hostable via Docker Compose or Helm, so your telemetry and prompts never have to leave your infrastructure.
No. OpenLIT’s SDK auto-instruments 90+ LLMs, agent frameworks, and vector databases with zero code changes, or you can call openlit.init() once for manual instrumentation if you want more control - both produce the same OpenTelemetry traces.
OpenLIT combines AI observability, evaluation, cost tracking, and prompt management in one open-source platform, instead of requiring separate tools for tracing, evals, and prompt versioning.