Armory
Source
Browse
Evals

wandb-weave-evals

Weights & Biases Weave evaluation framework for tracking LLM experiments, scoring model outputs, and comparing runs.

Score
97.9233 signals
Evidence
1,130 stars · 170 forks · 3 mentions
Last commit
as last read from GitHub; most reads are from 2 Sep 2026 or later
Listed

Install

No one-command install. Set it up from its source.

Alternatives · Evals

  1. trulens3,530 stars · 335 forks · 1 mention98.783
  2. braintrust27 stars · 12 forks82.319

What it is

Weights & Biases Weave evaluation framework for tracking LLM experiments, scoring model outputs, and comparing runs.

When to use it

Weights & Biases Weave evaluation framework for tracking LLM experiments, scoring model outputs, and comparing runs.

Notes

Curated evals entry. Verified 2026-05-27.