145 results in Memory, Identity, Evals, Observability · page 5 of 7
Harness Claude Code Cursor Codex Gemini OpenCode agntcy-identity Use when agents, MCP servers and multi-agent systems all need issued identities that another party can verify.
91.4 101 stars · 20 forks
Harness claude codex cursor gemini opencode
cp138-seed identity
vellum-evals Vellum evaluation SDK for running LLM test suites with custom metrics, dataset pinning, and CI workflow integration.
90.5 82 stars · 20 forks
Harness claude cursor codex opencode gemini
evals sdk ci dataset
character-card-spec-v3 Use when authoring an agent persona against the current character-card standard, with lorebooks, assets and decorators.
90.2 109 stars · 11 forks
Harness claude codex cursor gemini opencode
cp138-seed identity
agentrhq-authsome Use when agents must stay logged in to third-party services without ever seeing your credentials.
88.6 87 stars · 9 forks
Harness claude codex cursor gemini opencode
cp138-seed identity
gebruder-wirken Use when autonomous agents need one gateway that holds their credentials, isolates them per channel, and logs every session.
88.6 171 stars · 5 forks
Harness claude codex cursor gemini opencode
cp138-seed identity
opena2a-agent-identity-management Use when non-human identities need the same lifecycle a workforce IAM gives people — issue, authorize, audit, revoke.
88.5 59 stars · 18 forks
Harness claude codex cursor gemini opencode
cp138-seed identity
vestauth Use when an agent needs to authenticate as itself to services, without you hand-rolling keys and rotation.
87.5 166 stars · 4 forks
Harness claude codex cursor gemini opencode
cp138-seed identity
vorim-ai-labs-vorim-mcp-server Use when you want agent identity, scoped permissions and an audit trail exposed to the agent as tools it can call.
87.4 73 stars · 8 forks
Harness claude codex cursor gemini opencode
cp138-seed identity
deeptrail-deepsecure Use when agents move from prototype to production and their identities, secrets and access need managing centrally.
85.0 52 stars · 7 forks
Harness claude codex cursor gemini opencode
cp138-seed identity
agentcommunity-agent-identity-discovery Use when you want to find an agent's interface starting from nothing but a domain name.
84.8 47 stars · 8 forks
Harness claude codex cursor gemini opencode
cp138-seed identity
siddhant-k-code-agentic-authz Use when 'which agent may do what, to whose data' needs to be a policy decision rather than a code branch.
83.9 65 stars · 4 forks
Harness claude codex cursor gemini opencode
cp138-seed identity
kyros-ai Use when the agent's memory needs to resolve its own contradictions and forget on a schedule.
82.5 96 stars · 2 forks
Harness claude codex cursor gemini opencode
cp138-seed memory
vouch-protocol-vouch Use when agents acting in the world need identities that can be vouched for and held accountable afterwards.
82.5 27 stars · 13 forks
Harness claude codex cursor gemini opencode
cp138-seed identity
braintrust Developer platform for logging, evaluating, and comparing LLM experiments with dataset versioning and scoring functions.
82.3 27 stars · 12 forks
Harness claude cursor codex opencode gemini
evals experiment-tracking sdk
claudia-statusline High-performance Rust-based statusline for Claude Code with persistent stats tracking, progress bars, and optional cloud sync. Features SQLite-first persistence, git integration, context progress bars, burn rate calculation, XDG-compliant with theme support (dark/light, NO_COLOR).
81.2 36 stars · 5 forks
Harness claude
claude-code status-lines
auth0-auth0-ai-js Use when an agent must act on a user's behalf against a third-party API with the user's own consent and scope.
81.1 15 stars · 13 forks
Harness claude codex cursor gemini opencode
cp138-seed identity
newtype-ai-nit Use when you want an agent's identity to be versioned and inspectable the way a repository is.
80.5 128 stars · 1 fork
Harness claude codex cursor gemini opencode
cp138-seed identity
galileo-evaluate Galileo evaluation and observability SDK for detecting hallucinations, data errors, and model weaknesses in LLM pipelines.
80.3 22 stars · 11 forks
Harness claude cursor codex opencode gemini
evals hallucination observability sdk
alien-id-agent-id Use when each agent needs its own keypair so its actions can be signed and later attributed.
79.2 38 stars · 3 forks
Harness claude codex cursor gemini opencode
cp138-seed identity
mandarwagh9-machineauth Use when agents need to reach APIs and tools with their own credentials and permission scopes.
77.3 61 stars · 1 fork
Harness claude codex cursor gemini opencode
cp138-seed identity
authsec-ai Use when agent-to-agent calls need mutual authentication and a place to keep the agent's own secrets.
75.1 16 stars · 6 forks
Harness claude codex cursor gemini opencode
cp138-seed identity
honeycomb Honeycomb's OpenTelemetry-native observability platform: a high-cardinality event store ideal for tracing LLM pipelines and debugging slow agent traces.
75.1 16 stars · 6 forks
Harness claude cursor codex opencode gemini
observability opentelemetry tracing
baserun Baserun captures LLM traces via a lightweight decorator-based SDK and provides a dashboard for debugging prompt chains, testing variants, and measuring quality.
74.4 16 stars · 5 forks
Harness claude cursor codex opencode gemini
observability tracing testing
soth-ai-agentfacts-py Use when an agent should publish verifiable facts about itself — capabilities, provenance, operator — that another party can check.
73.3 33 stars · 1 fork
Harness claude codex cursor gemini opencode
cp138-seed identity
Previous Page 5 of 7 Next
Browse · Armory