245 results in Evals, Hooks, Memory, Identity · page 5 of 11
Harness Claude Code Cursor Codex Gemini OpenCode braintrust Developer platform for logging, evaluating, and comparing LLM experiments with dataset versioning and scoring functions.
82.3 27 stars · 12 forks
Harness claude cursor codex opencode gemini
evals experiment-tracking sdk
auth0-auth0-ai-js Use when an agent must act on a user's behalf against a third-party API with the user's own consent and scope.
81.1 15 stars · 13 forks
Harness claude codex cursor gemini opencode
cp138-seed identity
newtype-ai-nit Use when you want an agent's identity to be versioned and inspectable the way a repository is.
80.5 128 stars · 1 fork
Harness claude codex cursor gemini opencode
cp138-seed identity
galileo-evaluate Galileo evaluation and observability SDK for detecting hallucinations, data errors, and model weaknesses in LLM pipelines.
80.3 22 stars · 11 forks
Harness claude cursor codex opencode gemini
evals hallucination observability sdk
alien-id-agent-id Use when each agent needs its own keypair so its actions can be signed and later attributed.
79.2 38 stars · 3 forks
Harness claude codex cursor gemini opencode
cp138-seed identity
mandarwagh9-machineauth Use when agents need to reach APIs and tools with their own credentials and permission scopes.
77.3 61 stars · 1 fork
Harness claude codex cursor gemini opencode
cp138-seed identity
parry Prompt injection scanner for Claude Code hooks. Scans tool inputs and outputs for injection attacks, secrets, and data exfiltration attempts. In early development when it was listed.
75.4 45 stars · 1 fork
Harness claude
hook
authsec-ai Use when agent-to-agent calls need mutual authentication and a place to keep the agent's own secrets.
75.1 16 stars · 6 forks
Harness claude codex cursor gemini opencode
cp138-seed identity
britfix Claude outputs American spellings by default, which can have an impact on: professional credibility, compliance, documentation, and more. Britfix converts to British English, with a Claude Code hook for automatic conversion as files are written. Context-aware: handles code files intelligently by only converting comments and docstrings, never identifiers or string literals.
74.5 18 stars · 4 forks
Harness claude
claude-code hooks
soth-ai-agentfacts-py Use when an agent should publish verifiable facts about itself — capabilities, provenance, operator — that another party can check.
73.3 33 stars · 1 fork
Harness claude codex cursor gemini opencode
cp138-seed identity
clawsouls Use when you would rather install a ready-made agent persona than write one, and want it to work across OpenClaw, Claude Code and Cursor.
72.8 21 stars · 2 forks
Harness claude codex cursor gemini opencode
cp138-seed identity
kanoniv-agent-auth Use when authority must be handed down a chain of agents and each hop stays provable.
72.7 15 stars · 4 forks
Harness claude codex cursor gemini opencode
cp138-seed identity
clawsouls-soulspec Use when you want one file to define an agent's persistent identity in a way any compatible runtime can load.
69.7 21 stars · 1 fork
Harness claude codex cursor gemini opencode
cp138-seed identity
humanloop-evals Humanloop Python SDK with integrated evals, dataset versioning, and human + LLM judge scoring for production pipelines.
69.0 12 stars · 3 forks
Harness claude cursor codex opencode gemini
evals human-eval dataset sdk
intelliger-ai-oati Use when an agent's authority, and every action it took under that authority, must be verifiable after the fact.
64.8 13 stars · 1 fork
Harness claude codex cursor gemini opencode
cp138-seed identity
wikimem Use to give an agent a queryable wiki knowledge base when memory should be a navigable knowledge graph, not just a flat log: ingest files, folders, and URLs into a linked vault, then search or ask it in natural language.
63.4 7 stars · 4 forks
Harness claude
memory knowledge-base wiki ingest
chrisdbaldwin-masques Use when an agent needs to put on a temporary role — a bundle of intent, context and lens — for one task and take it off afterwards.
59.8 14 stars
Harness claude codex cursor gemini opencode
cp138-seed identity
imphillip-soultavern Use when you have character cards from the roleplay ecosystem and want them as SOUL.md personas an agent runtime can load.
59.0 13 stars
Harness claude codex cursor gemini opencode
cp138-seed identity
sunilp-aip Use when an agent's identity must carry across both MCP and agent-to-agent calls under one delegable scheme.
58.1 6 stars · 2 forks
Harness claude codex cursor gemini opencode
cp138-seed identity
amirf194-wingfoot Use when your agent keeps getting blocked by sites and you need it to present a verifiable bot identity and be told why it failed.
57.2 7 stars · 1 fork
Harness claude codex cursor gemini opencode
cp138-seed identity
paymanai-sigilum Use when an agent's identity has to leave an auditable trail rather than just gate a request.
54.8 9 stars
Harness claude codex cursor gemini opencode
cp138-seed identity
omnidotdev-persona-json Use when a non-human actor needs a portable, machine-readable identity document that travels with it between systems.
38.4 3 stars
Harness claude codex cursor gemini opencode
cp138-seed identity
agentbench Use to put a number on harness quality: run an agent harness against a task set and get a score, so harness changes are validated by evidence. It is the eval backbone of a self-improving loop.
Unranked No signals yet
Harness claude codex
eval benchmark scoring harness
agents-md-loader Automatically loads AGENTS.md configuration file content at session start to ensure Claude Code follows project-specific agent behavior. Only loads if AGENTS.md exists, otherwise passes empty context. Supports the universal AGENTS.md standard for cross-platform AI assistant compatibility.
Unranked No signals yet
Harness claude
automation hooks
Previous Page 5 of 11 Next
Browse · Armory