Armory
Source

Browse

Search and filter by type across the catalog

1,241 results in Identity, Evals, Skills, Observability · page 5 of 52

IdentityExperimental

vorim-ai-labs-vorim-mcp-server

Use when you want agent identity, scoped permissions and an audit trail exposed to the agent as tools it can call.

87.473 stars · 8 forks
Harnessclaudecodexcursorgeminiopencode
cp138-seedidentity
No one-command install · SourceDetails
IdentityExperimental

deeptrail-deepsecure

Use when agents move from prototype to production and their identities, secrets and access need managing centrally.

85.052 stars · 7 forks
Harnessclaudecodexcursorgeminiopencode
cp138-seedidentity
No one-command install · SourceDetails
IdentityExperimental

agentcommunity-agent-identity-discovery

Use when you want to find an agent's interface starting from nothing but a domain name.

84.847 stars · 8 forks
Harnessclaudecodexcursorgeminiopencode
cp138-seedidentity
No one-command install · SourceDetails
IdentityExperimental

siddhant-k-code-agentic-authz

Use when 'which agent may do what, to whose data' needs to be a policy decision rather than a code branch.

83.965 stars · 4 forks
Harnessclaudecodexcursorgeminiopencode
cp138-seedidentity
No one-command install · SourceDetails
IdentityExperimental

vouch-protocol-vouch

Use when agents acting in the world need identities that can be vouched for and held accountable afterwards.

82.527 stars · 13 forks
Harnessclaudecodexcursorgeminiopencode
cp138-seedidentity
No one-command install · SourceDetails
EvalsPreview

braintrust

Developer platform for logging, evaluating, and comparing LLM experiments with dataset versioning and scoring functions.

82.327 stars · 12 forks
Harnessclaudecursorcodexopencodegemini
evalsexperiment-trackingsdk
No one-command install · SourceDetails
ObservabilityPreview

claudia-statusline

High-performance Rust-based statusline for Claude Code with persistent stats tracking, progress bars, and optional cloud sync. Features SQLite-first persistence, git integration, context progress bars, burn rate calculation, XDG-compliant with theme support (dark/light, NO_COLOR).

81.236 stars · 5 forks
Harnessclaude
claude-codestatus-lines
No one-command install · SourceDetails
IdentityExperimental

auth0-auth0-ai-js

Use when an agent must act on a user's behalf against a third-party API with the user's own consent and scope.

81.115 stars · 13 forks
Harnessclaudecodexcursorgeminiopencode
cp138-seedidentity
No one-command install · SourceDetails
IdentityExperimental

newtype-ai-nit

Use when you want an agent's identity to be versioned and inspectable the way a repository is.

80.5128 stars · 1 fork
Harnessclaudecodexcursorgeminiopencode
cp138-seedidentity
No one-command install · SourceDetails
EvalsPreview

galileo-evaluate

Galileo evaluation and observability SDK for detecting hallucinations, data errors, and model weaknesses in LLM pipelines.

80.322 stars · 11 forks
Harnessclaudecursorcodexopencodegemini
evalshallucinationobservabilitysdk
No one-command install · SourceDetails
IdentityExperimental

alien-id-agent-id

Use when each agent needs its own keypair so its actions can be signed and later attributed.

79.238 stars · 3 forks
Harnessclaudecodexcursorgeminiopencode
cp138-seedidentity
No one-command install · SourceDetails
IdentityExperimental

mandarwagh9-machineauth

Use when agents need to reach APIs and tools with their own credentials and permission scopes.

77.361 stars · 1 fork
Harnessclaudecodexcursorgeminiopencode
cp138-seedidentity
No one-command install · SourceDetails
IdentityExperimental

authsec-ai

Use when agent-to-agent calls need mutual authentication and a place to keep the agent's own secrets.

75.116 stars · 6 forks
Harnessclaudecodexcursorgeminiopencode
cp138-seedidentity
No one-command install · SourceDetails
ObservabilityPreview

honeycomb

Honeycomb's OpenTelemetry-native observability platform: a high-cardinality event store ideal for tracing LLM pipelines and debugging slow agent traces.

75.116 stars · 6 forks
Harnessclaudecursorcodexopencodegemini
observabilityopentelemetrytracing
No one-command install · SourceDetails
ObservabilityPreview

baserun

Baserun captures LLM traces via a lightweight decorator-based SDK and provides a dashboard for debugging prompt chains, testing variants, and measuring quality.

74.416 stars · 5 forks
Harnessclaudecursorcodexopencodegemini
observabilitytracingtesting
No one-command install · SourceDetails
SkillsPreview

claude-mountaineering-skills

Claude Code skill that automates mountain route research for North American peaks. Aggregates data from 10+ mountaineering sources like Mountaineers.org, PeakBagger.com and SummitPost.com to generate detailed route beta reports with weather, avalanche conditions, and trip reports.

73.333 stars · 1 fork
Harnessclaude
skill
No one-command install · SourceDetails
IdentityExperimental

soth-ai-agentfacts-py

Use when an agent should publish verifiable facts about itself — capabilities, provenance, operator — that another party can check.

73.333 stars · 1 fork
Harnessclaudecodexcursorgeminiopencode
cp138-seedidentity
No one-command install · SourceDetails
IdentityExperimental

clawsouls

Use when you would rather install a ready-made agent persona than write one, and want it to work across OpenClaw, Claude Code and Cursor.

72.821 stars · 2 forks
Harnessclaudecodexcursorgeminiopencode
cp138-seedidentity
No one-command install · SourceDetails
IdentityExperimental

kanoniv-agent-auth

Use when authority must be handed down a chain of agents and each hop stays provable.

72.715 stars · 4 forks
Harnessclaudecodexcursorgeminiopencode
cp138-seedidentity
No one-command install · SourceDetails
IdentityExperimental

clawsouls-soulspec

Use when you want one file to define an agent's persistent identity in a way any compatible runtime can load.

69.721 stars · 1 fork
Harnessclaudecodexcursorgeminiopencode
cp138-seedidentity
No one-command install · SourceDetails
EvalsPreview

humanloop-evals

Humanloop Python SDK with integrated evals, dataset versioning, and human + LLM judge scoring for production pipelines.

69.012 stars · 3 forks
Harnessclaudecursorcodexopencodegemini
evalshuman-evaldatasetsdk
No one-command install · SourceDetails
SkillsPreview

read-only-postgres

Read-only PostgreSQL query skill for Claude Code. Executes SELECT/SHOW/EXPLAIN/WITH queries across configured databases with strict validation, timeouts, and row limits. Supports multiple connections with descriptions for database selection.

65.614 stars · 1 fork
Harnessclaude
claude-codeagent-skills
No one-command install · SourceDetails
IdentityExperimental

intelliger-ai-oati

Use when an agent's authority, and every action it took under that authority, must be verifiable after the fact.

64.813 stars · 1 fork
Harnessclaudecodexcursorgeminiopencode
cp138-seedidentity
No one-command install · SourceDetails
IdentityExperimental

chrisdbaldwin-masques

Use when an agent needs to put on a temporary role — a bundle of intent, context and lens — for one task and take it off afterwards.

59.814 stars
Harnessclaudecodexcursorgeminiopencode
cp138-seedidentity
No one-command install · SourceDetails
Browse · Armory