Armory
Source

Browse

Search and filter by type across the catalog

1,480 results in Workflows, Sub-Agents, Evals · page 4 of 62

WorkflowsPreview

ccoutputstyles

CLI tool and template gallery for customizing Claude Code output styles with pre-built templates. Features over 15 templates at the time of writing!

82.955 stars · 4 forks
Harnessclaude
output-style
No one-command install · SourceDetails
WorkflowsPreview

claude-code-agent-teams-exercises

Practical exercises for Claude Code Agent Teams - 6 exercises + 2 capstones covering team creation, task coordination, quality hooks, and parallel code review - good learning resource.

82.632 stars · 9 forks
Harnessclaude
claude-codeworkflows-knowledge-guides
No one-command install · SourceDetails
WorkflowsExperimental

ai-mocks

Mock HTTP/SSE and LLM servers, inspired by wiremock, with response streaming and SSE

82.652 stars · 4 forks
Harnessclaudecursorcodexopencodegemini
a2aagent-to-agentframeworks
No one-command install · SourceDetails
EvalsPreview

braintrust

Developer platform for logging, evaluating, and comparing LLM experiments with dataset versioning and scoring functions.

82.327 stars · 12 forks
Harnessclaudecursorcodexopencodegemini
evalsexperiment-trackingsdk
No one-command install · SourceDetails
WorkflowsPreview

awesome-claude-code-output-styles-that-i-really-like

A fun and moderately amusing collection of experimental output styles.

81.678 stars · 2 forks
Harnessclaude
output-style
No one-command install · SourceDetails
WorkflowsPreview

laravel-tall-stack-ai-development-starter-kit

Transform your Laravel TALL (Tailwind, AlpineJS, Laravel, Livewire) stack development with comprehensive Claude Code configurations that provide intelligent assistance, systematic workflows, and domain expert consultation.

81.041 stars · 4 forks
Harnessclaude
claude-codeworkflows-knowledge-guides
No one-command install · SourceDetails
WorkflowsPreview

linux-desktop-slash-commands

A library of slash commands intended specifically to facilitate common and advanced operations on Linux desktop environments (although many would also be useful on Linux servers). Command groups include hardware benchmarking, filesystem organisation, and security posture validation.

80.630 stars · 6 forks
Harnessclaude
claude-codeslash-commands
No one-command install · SourceDetails
EvalsPreview

galileo-evaluate

Galileo evaluation and observability SDK for detecting hallucinations, data errors, and model weaknesses in LLM pipelines.

80.322 stars · 11 forks
Harnessclaudecursorcodexopencodegemini
evalshallucinationobservabilitysdk
No one-command install · SourceDetails
WorkflowsExperimental

azure-openai-llm-cookbook

A one-stop hub with 100+ Azure OpenAI sample code organized by topic for quick reference

72.016 stars · 3 forks
Harnessclaudecursorcodexopencodegemini
a2aagent-to-agenttutorials-learning-resources
No one-command install · SourceDetails
EvalsPreview

humanloop-evals

Humanloop Python SDK with integrated evals, dataset versioning, and human + LLM judge scoring for production pipelines.

69.012 stars · 3 forks
Harnessclaudecursorcodexopencodegemini
evalshuman-evaldatasetsdk
No one-command install · SourceDetails
WorkflowsExperimental

a2a-walk-through

A2A Concept walkthrough

60.615 stars
Harnessclaudecursorcodexopencodegemini
a2aagent-to-agenttutorials-learning-resources
No one-command install · SourceDetails
WorkflowsExperimental

agent-rank

Documentation about agent ranking

60.69 stars · 1 fork
Harnessclaudecursorcodexopencodegemini
a2aagent-to-agentdocumentation
No one-command install · SourceDetails
WorkflowsExperimental

qredence-agentic-kernel

A flexible foundation AI system for creating A2A-compatible autonomous AI agents

59.814 stars
Harnessclaudecursorcodexopencodegemini
a2aagent-to-agentframeworks
No one-command install · SourceDetails
WorkflowsPreview

gen-alpha-slang

An output style that makes Claude Code write in Gen Alpha slang, listed for its humour.

59.08 stars · 1 fork
Harnessclaude
output-style
No one-command install · SourceDetails
WorkflowsPreview

ralph-wiggum-bdd

A standalone Bash script for Behavior-Driven Development with Ralph Wiggum Loop. In principle, while running unattended, the script can keep code and requirements in sync, but in practice it still requires interactive human supervision, so it supports both modes. The script is standalone and can be modified and committed into your project.

53.18 stars
Harnessclaude
claude-codeworkflows-knowledge-guides
No one-command install · SourceDetails
WorkflowsExperimental

a2a-docs-zh

Agent2Agent Protocol (A2A) Chinese documentation

53.18 stars
Harnessclaudecursorcodexopencodegemini
a2aagent-to-agentdocumentation
No one-command install · SourceDetails
WorkflowsExperimental

llm-tutorials

AI drawing learning guide

38.43 stars
Harnessclaudecursorcodexopencodegemini
a2aagent-to-agenttutorials-learning-resources
No one-command install · SourceDetails
WorkflowsExperimental

mcp-o

MCP-based orchestration framework

38.43 stars
Harnessclaudecursorcodexopencodegemini
a2aagent-to-agentframeworks
No one-command install · SourceDetails
WorkflowsPreview

claude-code-output-styles-debugging

A small set of output styles for debugging: root cause analysis and a systematic, methodical approach to fixing bugs in Claude Code.

38.02 stars · 1 fork
Harnessclaude
claude-codeoutput-styles
No one-command install · SourceDetails
WorkflowsExperimental

agent-hub

Keeps AI agents discoverable, composable, and resilient

29.01 star · 1 fork
Harnessclaudecursorcodexopencodegemini
a2aagent-to-agentframeworks
No one-command install · SourceDetails
WorkflowsExperimental

a2a-demo

A2A Demo implementation

22.51 star
Harnessclaudecursorcodexopencodegemini
a2aagent-to-agenttutorials-learning-resources
No one-command install · SourceDetails
WorkflowsExperimental

a2a-zh

Chinese documentation for A2A protocol

22.51 star
Harnessclaudecursorcodexopencodegemini
a2aagent-to-agentdocumentation
No one-command install · SourceDetails
EvalsPreview

agentbench

Use to put a number on harness quality: run an agent harness against a task set and get a score, so harness changes are validated by evidence. It is the eval backbone of a self-improving loop.

UnrankedNo signals yet
Harnessclaudecodex
evalbenchmarkscoringharness
No one-command install · SourceDetails
WorkflowsExperimental

a2a-example

Example implementation of A2A protocol

UnrankedNo signals yet
Harnessclaudecursorcodexopencodegemini
a2aagent-to-agenttutorials-learning-resources
No one-command install · SourceDetails
Browse · Armory