wandb-weave-evals
Weights & Biases Weave evaluation framework for tracking LLM experiments, scoring model outputs, and comparing runs.
- Score
- 97.9233 signals
- Evidence
- 1,130 stars · 170 forks · 3 mentions
- Last commit
- as last read from GitHub; most reads are from 2 Sep 2026 or later
- Listed
Install
No one-command install. Set it up from its source.
Alternatives · Evals
- trulens3,530 stars · 335 forks · 1 mention98.783
- braintrust27 stars · 12 forks82.319
What it is
Weights & Biases Weave evaluation framework for tracking LLM experiments, scoring model outputs, and comparing runs.
When to use it
Weights & Biases Weave evaluation framework for tracking LLM experiments, scoring model outputs, and comparing runs.
Notes
Curated evals entry. Verified 2026-05-27.