Phoenix

AI observability and evaluation from Arize

Open SourceSelf Hosted
0.0 (0)

About

Phoenix from Arize is an open-source observability and evaluation platform for LLM and agent applications. It captures OpenTelemetry-style traces, runs evaluations against the captured data, and visualizes prompt-response pairs, embeddings, and metrics in a local UI. Auto-instrumentation covers OpenAI, Anthropic, Google GenAI, Bedrock, LangGraph, Vercel AI SDK, CrewAI, LlamaIndex, DSPy, and others. Self-host via Docker or use Arize's hosted edition.

Reviews (0)

Leave a Review

No reviews yet. Be the first to review!

Details

Price
Free
Platform
Hybrid
Difficulty
Easy (2/5)
License
Elastic-2.0
Added
Jan 29, 2026

Related Tools

UK AI Security Institute framework for large language model evaluations and benchmarks.

Open SourceSelf HostedOffline
Easy
Featured

ML experiment tracking, visualization, and collaboration

Open Source
Easy
Featured

Open source LLM engineering platform for tracing and analytics

Open SourceSelf Hosted
Easy

Open-source library for evaluating and tracking LLM applications.

Open SourceSelf Hosted
Easy

Open-source AI observability platform for tracing, evaluation, and experimentation.

Open SourceSelf Hosted
Easy

Python framework for unit testing and evaluating LLM applications with metrics like G-Eval.

Open SourceSelf HostedOffline
Easy
Browse all AI Observability & Evaluation tools