humanloop-evals
Humanloop Python SDK with integrated evals, dataset versioning, and human + LLM judge scoring for production pipelines.
- Score
- 69.0002 signals
- Evidence
- 12 stars · 3 forks
- Last commit
- as last read from GitHub; most reads are from 2 Sep 2026 or later
- Listed
Install
No one-command install. Set it up from its source.
Alternatives · Evals
- braintrust27 stars · 12 forks82.319
What it is
Humanloop Python SDK with integrated evals, dataset versioning, and human + LLM judge scoring for production pipelines.
When to use it
Humanloop Python SDK with integrated evals, dataset versioning, and human + LLM judge scoring for production pipelines.
Notes
Curated evals entry. Verified 2026-05-27.