164 results in Observability, Evals, Infrastructure, Identity · page 5 of 7
Harness Claude Code Cursor Codex Gemini OpenCode vorim-ai-labs-vorim-mcp-server Use when you want agent identity, scoped permissions and an audit trail exposed to the agent as tools it can call.
87.4 73 stars · 8 forks
Harness claude codex cursor gemini opencode
cp138-seed identity
deeptrail-deepsecure Use when agents move from prototype to production and their identities, secrets and access need managing centrally.
85.0 52 stars · 7 forks
Harness claude codex cursor gemini opencode
cp138-seed identity
agentcommunity-agent-identity-discovery Use when you want to find an agent's interface starting from nothing but a domain name.
84.8 47 stars · 8 forks
Harness claude codex cursor gemini opencode
cp138-seed identity
siddhant-k-code-agentic-authz Use when 'which agent may do what, to whose data' needs to be a policy decision rather than a code branch.
83.9 65 stars · 4 forks
Harness claude codex cursor gemini opencode
cp138-seed identity
vouch-protocol-vouch Use when agents acting in the world need identities that can be vouched for and held accountable afterwards.
82.5 27 stars · 13 forks
Harness claude codex cursor gemini opencode
cp138-seed identity
braintrust Developer platform for logging, evaluating, and comparing LLM experiments with dataset versioning and scoring functions.
82.3 27 stars · 12 forks
Harness claude cursor codex opencode gemini
evals experiment-tracking sdk
claudia-statusline High-performance Rust-based statusline for Claude Code with persistent stats tracking, progress bars, and optional cloud sync. Features SQLite-first persistence, git integration, context progress bars, burn rate calculation, XDG-compliant with theme support (dark/light, NO_COLOR).
81.2 36 stars · 5 forks
Harness claude
claude-code status-lines
auth0-auth0-ai-js Use when an agent must act on a user's behalf against a third-party API with the user's own consent and scope.
81.1 15 stars · 13 forks
Harness claude codex cursor gemini opencode
cp138-seed identity
newtype-ai-nit Use when you want an agent's identity to be versioned and inspectable the way a repository is.
80.5 128 stars · 1 fork
Harness claude codex cursor gemini opencode
cp138-seed identity
galileo-evaluate Galileo evaluation and observability SDK for detecting hallucinations, data errors, and model weaknesses in LLM pipelines.
80.3 22 stars · 11 forks
Harness claude cursor codex opencode gemini
evals hallucination observability sdk
e2b Open-source secure cloud sandboxes (Firecracker microVMs) for running AI-generated code. ~150ms cold start.
80.0 passed install test
Harness claude cursor codex opencode gemini
infrastructure sandbox
modal Serverless cloud platform for running Python functions, containers, and AI workloads with zero infra management.
80.0 passed install test
Harness claude cursor codex opencode gemini
infrastructure serverless
railway Zero-config cloud platform for deploying agent backends, databases, and services from a Git push.
80.0 passed install test
Harness claude cursor codex opencode gemini
infrastructure deploy
vercel Frontend cloud platform with serverless functions and AI SDK integrations for deploying agent-facing UIs.
80.0 passed install test
Harness claude cursor codex opencode gemini
infrastructure deploy
alien-id-agent-id Use when each agent needs its own keypair so its actions can be signed and later attributed.
79.2 38 stars · 3 forks
Harness claude codex cursor gemini opencode
cp138-seed identity
mandarwagh9-machineauth Use when agents need to reach APIs and tools with their own credentials and permission scopes.
77.3 61 stars · 1 fork
Harness claude codex cursor gemini opencode
cp138-seed identity
authsec-ai Use when agent-to-agent calls need mutual authentication and a place to keep the agent's own secrets.
75.1 16 stars · 6 forks
Harness claude codex cursor gemini opencode
cp138-seed identity
honeycomb Honeycomb's OpenTelemetry-native observability platform: a high-cardinality event store ideal for tracing LLM pipelines and debugging slow agent traces.
75.1 16 stars · 6 forks
Harness claude cursor codex opencode gemini
observability opentelemetry tracing
baserun Baserun captures LLM traces via a lightweight decorator-based SDK and provides a dashboard for debugging prompt chains, testing variants, and measuring quality.
74.4 16 stars · 5 forks
Harness claude cursor codex opencode gemini
observability tracing testing
soth-ai-agentfacts-py Use when an agent should publish verifiable facts about itself — capabilities, provenance, operator — that another party can check.
73.3 33 stars · 1 fork
Harness claude codex cursor gemini opencode
cp138-seed identity
clawsouls Use when you would rather install a ready-made agent persona than write one, and want it to work across OpenClaw, Claude Code and Cursor.
72.8 21 stars · 2 forks
Harness claude codex cursor gemini opencode
cp138-seed identity
kanoniv-agent-auth Use when authority must be handed down a chain of agents and each hop stays provable.
72.7 15 stars · 4 forks
Harness claude codex cursor gemini opencode
cp138-seed identity
clawsouls-soulspec Use when you want one file to define an agent's persistent identity in a way any compatible runtime can load.
69.7 21 stars · 1 fork
Harness claude codex cursor gemini opencode
cp138-seed identity
humanloop-evals Humanloop Python SDK with integrated evals, dataset versioning, and human + LLM judge scoring for production pipelines.
69.0 12 stars · 3 forks
Harness claude cursor codex opencode gemini
evals human-eval dataset sdk
Previous Page 5 of 7 Next
Browse · Armory