Weights & Biases
ML experiment tracking, LLM observability, and evaluation platform for AI teams
Weights & Biases provides experiment tracking and LLM observability tools for AI developers. W&B Models lets you log experiments, track hyperparameters, visualize training metrics, compare runs, and manage model artifacts across all major ML frameworks. W&B Weave adds tracing, evaluation, and monitoring for LLM applications, automatically logging inputs, outputs, and metadata with support for multi-modal tracking. The platform includes a generous free tier with unlimited tracked hours and projects, with Pro and Enterprise plans for larger teams.
Pricing: Free / monthly subscriptions
Weights & Biases Alternatives
Explore 55 products in the Observability & Analytics category. View all Weights & Biases alternatives.
Klu
Collaborate on prompts, evaluate, and optimize LLM-powered Apps with Klu.
Langfuse
Traces, evals, prompt management and metrics to debug and improve your LLM application.
Agenta
Open-source prompt management, evaluation, and observability for LLM apps
Arize AI
AI observability platform with tracing, evaluation, and monitoring for LLM and ML applications
Portkey
AI gateway for routing to 1,600+ LLMs with observability, guardrails, and prompt management
DeepEval
Open-source LLM evaluation framework with 50+ metrics for testing agents, RAG, and chatbots
Work on Weights & Biases? Feature it at the top of Observability & Analytics.
Is your product missing?