Armory
Source

Browse

Search and filter by type across the catalog

845 results in Infrastructure, Workflows, Evals, Observability · page 5 of 36

WorkflowsPreview

simone

A broader project management workflow for Claude Code that encompasses not just a set of commands, but a system of documents, guidelines, and processes to facilitate project planning and execution.

96.4558 stars · 46 forks
Harnessclaude
claude-codeworkflows-knowledge-guides
No one-command install · SourceDetails
EvalsPreview

continuous-eval

Relari's modular evaluation library for LLM pipelines with deterministic + LLM-based metrics for RAG and agent workflows.

96.2517 stars · 38 forks
Harnessclaudecursorcodexopencodegemini
evalsragagentsmetrics
No one-command install · SourceDetails
ObservabilityPreview

claude-code-statusline

Enhanced 4-line statusline for Claude Code with themes, cost tracking, and MCP server monitoring

95.9476 stars · 35 forks
Harnessclaude
statuslineobservability
No one-command install · SourceDetails
ObservabilityPreview

phospho

Phospho is a text analytics and evaluation platform for LLM apps. It logs sessions, runs clustering, detects failures, and surfaces actionable insights.

95.8439 stars · 35 forks
Harnessclaudecursorcodexopencodegemini
observabilityanalyticsevals
No one-command install · SourceDetails
WorkflowsPreview

learn-faster-kit

An educational framework for Claude Code based on the "FASTER" approach to self-teaching, with agents, slash commands and tools for learning at your own pace through active learning and spaced repetition.

95.8380 stars · 41 forks
Harnessclaude
claude-codeworkflows-knowledge-guides
No one-command install · SourceDetails
WorkflowsPreview

agentic-workflow-patterns

A collection of agentic patterns from Anthropic's docs, each with a Mermaid diagram and a code example: sub-agent orchestration, progressive skills, parallel tool calling, master-clone architecture, wizard workflows and more. Also works with other providers.

95.2302 stars · 33 forks
Harnessclaude
claude-codeworkflows-knowledge-guides
No one-command install · SourceDetails
WorkflowsExperimental

deepseek-implementation

DeepSeek application development course companion code

94.8160 stars · 67 forks
Harnessclaudecursorcodexopencodegemini
a2aagent-to-agenttutorials-learning-resources
No one-command install · SourceDetails
WorkflowsExperimental

a2a-samples

Samples for A2A implementation

94.6107 stars · 70 forks
Harnessclaudecursorcodexopencodegemini
a2aagent-to-agenttutorials-learning-resources
No one-command install · SourceDetails
ObservabilityPreview

athina-ai

Athina AI provides developer-focused LLM monitoring and eval framework: real-time inference logging, automated evals, and regression detection in CI.

94.6301 stars · 23 forks
Harnessclaudecursorcodexopencodegemini
observabilityevalslogging
No one-command install · SourceDetails
WorkflowsExperimental

mentis

A powerful multi-agent orchestration framework built on LangGraph

94.5296 stars · 23 forks
Harnessclaudecursorcodexopencodegemini
a2aagent-to-agentframeworks
No one-command install · SourceDetails
EvalsPreview

metr-task-standard

METR's Task Standard: a specification and scaffold for creating agentic tasks used in autonomous agent capability evaluations.

94.3192 stars · 37 forks
Harnessclaudecursorcodexopencodegemini
evalsagentstask-standardsafety
No one-command install · SourceDetails
EvalsExperimental

harbor-framework-terminal-bench-2-1

Terminal-Bench 2.1

Contributed by Sentinel

94.3119 stars · 62 forks · 3 mentions
Harnessclaudecodexcursorgeminiopencode
evals
No one-command install · SourceDetails
ObservabilityPreview

claude-pace

A lightweight Bash + jq statusline for Claude Code that displays rate limit pace delta (burn rate vs. time remaining), 5h/7d usage percentage, context window usage, git branch and diff stats. Compares current consumption rate against time remaining in each rate limit window to indicate whether quota is being used faster or slower than the window allows. Single file with no external dependencies beyond jq.

93.6229 stars · 19 forks
Harnessclaude
claude-codestatus-lines
No one-command install · SourceDetails
WorkflowsExperimental

claude-mcp

Claude Unified Model Context Interaction Protocol

92.848 stars · 56 forks
Harnessclaudecursorcodexopencodegemini
a2aagent-to-agentframeworks
No one-command install · SourceDetails
WorkflowsPreview

ab-method

A principled, spec-driven workflow that transforms large problems into focused, incremental missions using Claude Code's specialized sub agents. Includes slash-commands, sub agents, and specialized workflows designed for specific parts of the SDLC.

92.5189 stars · 14 forks
Harnessclaude
claude-codeworkflows-knowledge-guides
No one-command install · SourceDetails
WorkflowsExperimental

a2a-mcp-tutorial

A tutorial on how to use Model Context Protocol by Anthropic and Agent2Agent Protocol by Google

92.4114 stars · 29 forks
Harnessclaudecursorcodexopencodegemini
a2aagent-to-agenttutorials-learning-resources
No one-command install · SourceDetails
WorkflowsExperimental

mcpyats

VibeOps - Cisco pyATS MCP Plus Many Other MCPs

90.669 stars · 30 forks
Harnessclaudecursorcodexopencodegemini
a2aagent-to-agentframeworks
No one-command install · SourceDetails
EvalsPreview

vellum-evals

Vellum evaluation SDK for running LLM test suites with custom metrics, dataset pinning, and CI workflow integration.

90.582 stars · 20 forks
Harnessclaudecursorcodexopencodegemini
evalssdkcidataset
No one-command install · SourceDetails
WorkflowsPreview

riper-workflow

Structured development workflow enforcing separation between Research, Innovate, Plan, Execute, and Review phases. Features consolidated subagents for context-efficiency, branch-aware memory bank, and strict mode enforcement for guided development.

89.693 stars · 11 forks
Harnessclaude
workflowguide
No one-command install · SourceDetails
WorkflowsExperimental

a2a-adk-mcp-tutorial

Multi-Agent Systems with Google's Agent Development Kit + A2A + MCP

88.258 stars · 16 forks
Harnessclaudecursorcodexopencodegemini
a2aagent-to-agenttutorials-learning-resources
No one-command install · SourceDetails
WorkflowsPreview

prd-generator

A Claude Code plugin that generates comprehensive Product Requirements Documents (PRDs) from conversation context. Invoke `/create-prd` after discussing requirements and it produces a complete PRD with all standard sections including Executive Summary, User Stories, MVP Scope, Architecture, Success Criteria, and Implementation Phases.

85.552 stars · 8 forks
Harnessclaude
claude-codeslash-commands
No one-command install · SourceDetails
WorkflowsExperimental

techie-talks-ai

Techie Talks AI Youtube Channel code repository

84.324 stars · 16 forks
Harnessclaudecursorcodexopencodegemini
a2aagent-to-agenttutorials-learning-resources
No one-command install · SourceDetails
WorkflowsExperimental

artinet-sdk

A JS/TS SDK for the Agent2Agent Protocol

84.246 stars · 7 forks
Harnessclaudecursorcodexopencodegemini
a2aagent-to-agentframeworks
No one-command install · SourceDetails
WorkflowsPreview

claude-code-docs

A mirror of the Anthropic© PBC documentation site for Claude/Code, but with bonus features like full-text search and query-time updates - a nice companion to `claude-code-docs` for up-to-the-minute, fully-indexed information so that Claude Code can read about itself.

83.753 stars · 5 forks
Harnessclaude
workflowguide
No one-command install · SourceDetails
Browse · Armory