Weights & Biases
ML experiment tracking, visualization, and collaboration
About
Weights and Biases is a platform for experiment tracking, dataset versioning, model registry, and visualization across machine learning pipelines. The wandb Python library logs metrics, hyperparameters, gradients, and artifacts from training runs to a hosted or self-hosted server with a web UI for comparisons and reports. The team also ships Weave, a companion library focused on tracing, evaluation, and monitoring of LLM and agent apps.
Reviews (0)
Leave a Review
No reviews yet. Be the first to review!
Details
- Category
- AI Observability & Evaluation
- Price
- Freemium
- Platform
- Hybrid
- Difficulty
- Easy (2/5)
- License
- MIT
- Added
- Jan 29, 2026
Related Tools
UK AI Security Institute framework for large language model evaluations and benchmarks.
Open source LLM engineering platform for tracing and analytics
Open-source library for evaluating and tracking LLM applications.
Open-source AI observability platform for tracing, evaluation, and experimentation.
SDK for monitoring AI agents with session replays, cost tracking, and OpenTelemetry export.
Python framework for unit testing and evaluating LLM applications with metrics like G-Eval.