Armory
Source

Browse

Search and filter by type across the catalog

1,169 results in Skills, Evals · page 4 of 49

SkillsPreview

address-github-comments

Use when you need to address review or issue comments on an open GitHub Pull Request using the gh CLI.

UnrankedNo signals yet
Harnessclaude
address-github-commentsskills
Details
SkillsPreview

aeon

This skill should be used for time series machine learning tasks including classification, regression, clustering, forecasting, anomaly detection, segmentation, and similarity search. Use when working with temporal data, sequential patterns, or time-indexed observations requiring specialized algorithms beyond standard ML approaches. Particularly suited for univariate and multivariate time series analysis with scikit-learn compatible APIs.

UnrankedNo signals yet
Harnessclaude
aeonskills
Details
SkillsPreview

agent-architecture-audit

Full-stack diagnostic for agent and LLM applications. Audits the 12-layer agent stack for wrapper regression, memory pollution, tool discipline failures, hidden repair loops, and rendering corruption. Produces severity-ranked findings with code-first fixes. Essential for developers building agent…

UnrankedNo signals yet
Harnessclaudecodexcursorgeminiopencode
skill
Details
SkillsPreview

agent-development

This skill should be used when the user asks to "create an agent", "add an agent", "write a subagent", "agent frontmatter", "when to use description", "agent examples", "agent tools", "agent colors", "autonomous agent", or needs guidance on agent structure, system prompts, triggering conditions, or agent development best practices for Claude Code plugins.

UnrankedNo signals yet
Harnessclaude
agent-developmentskills
Details
SkillsPreview

agent-eval

Head-to-head comparison of coding agents (Claude Code, Aider, Codex, etc.) on custom tasks with pass rate, cost, time, and consistency metrics

UnrankedNo signals yet
Harnessclaudecodexcursorgeminiopencode
agentsevalbenchmark
Details
SkillsPreview

agent-evaluation

Testing and benchmarking LLM agents including behavioral testing, capability assessment, reliability metrics, and production monitoring—where even top agents achieve less than 50% on real-world benchmarks Use when: agent testing, agent evaluation, benchmark agents, agent reliability, test agent.

UnrankedNo signals yet
Harnessclaude
agent-evaluationskills
Details
SkillsPreview

agent-harness-construction

Design and optimize AI agent action spaces, tool definitions, and observation formatting for higher completion rates.

UnrankedNo signals yet
Harnessclaudecodexcursorgeminiopencode
harnessagentstools
Details
SkillsPreview

agent-introspection-debugging

Structured self-debugging workflow for AI agent failures using capture, diagnosis, contained recovery, and introspection reports.

UnrankedNo signals yet
Harnessclaudecodexcursorgeminiopencode
skill
Details
SkillsPreview

agent-management

Create, manage, and orchestrate AI agents using the AI Maestro CLI. Use when the user asks to "create agent", "list agents", "delete agent", "hibernate agent", "wake agent", "install plugin", "show agent", "restart agent", or any agent lifecycle management task.

UnrankedNo signals yet
Harnessclaude
agent-managementskills
Details
SkillsPreview

agent-manager-skill

Manage multiple local CLI agents via tmux sessions (start/stop/monitor/assign) with cron-friendly scheduling.

UnrankedNo signals yet
Harnessclaude
agent-manager-skillskills
Details
SkillsPreview

agent-md-refactor

Refactor bloated AGENTS.md, CLAUDE.md, or similar agent instruction files to follow progressive disclosure principles. Splits monolithic files into organized, linked documentation.

UnrankedNo signals yet
Harnessclaude
agent-md-refactorskills
Details
SkillsPreview

agent-memory-mcp

A hybrid memory system that provides persistent, searchable knowledge management for AI agents (Architecture, Patterns, Decisions).

UnrankedNo signals yet
Harnessclaude
agent-memory-mcpskills
Details
SkillsPreview

agent-memory-systems

Memory is the cornerstone of intelligent agents. Without it, every interaction starts from zero. This skill covers the architecture of agent memory: short-term (context window), long-term (vector stores), and the cognitive architectures that organize them. Key insight: Memory isn't just storage - it's retrieval. A million stored facts mean nothing if you can't find the right one. Chunking, embedding, and retrieval strategies determine whether your agent remembers or forgets. The field is fragm

UnrankedNo signals yet
Harnessclaude
agent-memory-systemsskills
Details
SkillsPreview

agent-messaging

Send and receive cryptographically signed messages between AI agents using the Agent Messaging Protocol (AMP). Use when the user asks to "send a message to an agent", "check agent inbox", "message another agent", "reply to a message", "notify an agent", or any inter-agent communication task.

UnrankedNo signals yet
Harnessclaude
agent-messagingskills
Details
SkillsPreview

agent-payment-x402

Add x402 payment execution to AI agents with per-task budgets, spending controls, and non-custodial wallets. Supports Base through agentwallet-sdk and X Layer through OKX Payments / OKX Agent Payments Protocol.

UnrankedNo signals yet
Harnessclaudecodexcursorgeminiopencode
skill
Details
SkillsPreview

agent-sort

Build an evidence-backed ECC install plan for a specific repo by sorting skills, commands, rules, hooks, and extras into DAILY vs LIBRARY buckets using parallel repo-aware review passes. Use when ECC should be trimmed to what a project actually needs instead of loading the full bundle.

UnrankedNo signals yet
Harnessclaudecodexcursorgeminiopencode
skill
Details
SkillsPreview

agent-tool-builder

Tools are how AI agents interact with the world. A well-designed tool is the difference between an agent that works and one that hallucinates, fails silently, or costs 10x more tokens than necessary. This skill covers tool design from schema to error handling. JSON Schema best practices, description writing that actually helps the LLM, validation, and the emerging MCP standard that's becoming the lingua franca for AI tools. Key insight: Tool descriptions are more important than tool implementa

UnrankedNo signals yet
Harnessclaude
agent-tool-builderskills
Details
SkillsPreview

agentic-engineering

Operate as an agentic engineer using eval-first execution, decomposition, and cost-aware model routing.

UnrankedNo signals yet
Harnessclaudecodexcursorgeminiopencode
agentsengineeringeval
Details
SkillsPreview

agentic-os

Build persistent multi-agent operating systems on Claude Code. Covers kernel architecture, specialist agents, slash commands, file-based memory, scheduled automation, and state management without external databases.

UnrankedNo signals yet
Harnessclaudecodexcursorgeminiopencode
skill
Details
SkillsPreview

agile-product-owner

Agile product ownership toolkit for Senior Product Owner including INVEST-compliant user story generation, sprint planning, backlog management, and velocity tracking. Use for story writing, sprint planning, stakeholder communication, and agile ceremonies.

UnrankedNo signals yet
Harnessclaude
agile-product-ownerskills
Details
SkillsPreview

agirails-agent-payments

You are a payments engineer for the AI agent economy. Your job is to onboard agents onto the

UnrankedNo signals yet
Harnessclaude
agirails-agent-paymentsskills
Details
SkillsPreview

ai-agent-ai-spy

A talk by members of the Signal Foundation on how AI agents built into operating systems can be used for surveillance, and why companies are building them. A YouTube video.

UnrankedNo signals yet
Harnessclaude
claude-codeagent-skills
No one-command install · SourceDetails
SkillsPreview

ai-agents-architect

Expert in designing and building autonomous AI agents. Masters tool use, memory systems, planning strategies, and multi-agent orchestration. Use when: build agent, AI agent, autonomous agent, tool use, function calling.

UnrankedNo signals yet
Harnessclaude
ai-agents-architectskills
Details
SkillsPreview

ai-first-engineering

Engineering operating model for teams where AI agents generate a large share of implementation output.

UnrankedNo signals yet
Harnessclaudecodexcursorgeminiopencode
engineeringprocessagents
Details
Browse · Armory