Armory
Source

Browse

Search and filter by type across the catalog

2,612 results in Sub-Agents, Skills, Workflows, Evals · page 2 of 109

SkillsExperimental

hardikpandya-stop-slop

A skill file for removing AI tells from prose

Contributed by Sentinel

99.316,882 stars · 1,223 forks · 3 mentions
Harnessclaudecodexcursorgeminiopencode
skills
No one-command install · SourceDetails
WorkflowsPreview

claude-code-system-prompts

All parts of Claude Code's system prompt, including builtin tool descriptions, sub agent prompts (Plan/Explore/Task), utility prompts (CLAUDE.md, compact, Bash cmd, security review, agent creation, etc.). Updated for each Claude Code version.

99.312,546 stars · 2,043 forks
Harnessclaude
workflowguide
No one-command install · SourceDetails
SkillsExperimental

nidhinjs-prompt-master

A Claude skill that writes the accurate prompts for any AI tool. Zero tokens or credits wasted. Full context and memory retention

Contributed by Sentinel

99.312,473 stars · 1,462 forks · 3 mentions
Harnessclaudecodexcursorgeminiopencode
skills
No one-command install · SourceDetails
WorkflowsExperimental

pocketflow

Pocket Flow: 100-line LLM framework that lets agents build agents

99.211,139 stars · 1,217 forks
Harnessclaudecursorcodexopencodegemini
a2aagent-to-agentframeworks
No one-command install · SourceDetails
EvalsPreview

phoenix

Arize Phoenix: open-source LLM observability with built-in evals, span tracing, and dataset curation for RAG and agents.

99.211,286 stars · 1,086 forks · 2 mentions
Harnessclaudecursorcodexopencodegemini
evalsobservabilityragagents
No one-command install · SourceDetails
SkillsPreview

fullstack-dev-skills

A Claude Code plugin with 65 skills for full-stack development across many frameworks, 9 workflow commands for Jira and Confluence, and a /common-ground command that lists Claude's assumptions about your project.

99.211,283 stars · 1,080 forks
Harnessclaude
claude-codeagent-skills
No one-command install · SourceDetails
WorkflowsPreview

claude-code-infrastructure-showcase

An approach to working with Skills that uses hooks to make Claude select and activate the right Skill for the current context. Documented, and adaptable to other projects and workflows.

99.210,016 stars · 1,231 forks
Harnessclaude
claude-codeworkflows-knowledge-guides
No one-command install · SourceDetails
SkillsExperimental

agricidaniel-claude-ads

Claude-first paid-media operations skill for Claude Code across 12 ad platforms (Google, Meta, YouTube, LinkedIn, TikTok, Microsoft, Apple, Amazon, Reddit, Pinterest, Snapchat, X): source-grounded audits, deterministic scoring, versioned JSON reports, and capability-gated account changes.

Contributed by Sentinel

99.28,670 stars · 1,293 forks · 3 mentions
Harnessclaudecodexcursorgeminiopencode
No one-command install · SourceDetails
WorkflowsPreview

harness

A meta-skill that designs domain-specific agent teams, defines specialized agents, and generates the skills they use. Resources are in Korean but can produce high-quality English-language output.

99.28,876 stars · 1,256 forks
Harnessclaude
workflowguideteams
No one-command install · SourceDetails
WorkflowsPreview

claude-code-tips

35+ short Claude Code tips covering voice input, system prompt patching, container workflows for risky tasks, conversation cloning, multi-model orchestration with Gemini CLI and more, with demos, working scripts and a plugin.

99.210,008 stars · 809 forks
Harnessclaude
claude-codeworkflows-knowledge-guides
No one-command install · SourceDetails
WorkflowsPreview

ralph-for-claude-code

An autonomous AI development framework that enables Claude Code to work iteratively on projects until completion. Features intelligent exit detection, rate limiting, circuit breaker patterns, and comprehensive safety guardrails to prevent infinite loops and API overuse. Built with Bash, integrated with tmux for live monitoring, and includes 75+ comprehensive tests.

99.19,613 stars · 721 forks
Harnessclaude
claude-codeworkflows-knowledge-guides
No one-command install · SourceDetails
WorkflowsPreview

claude-code-pm

A project-management workflow for Claude Code with specialised agents, slash commands and documentation.

99.18,362 stars · 837 forks
Harnessclaude
workflowguide
No one-command install · SourceDetails
EvalsPreview

swe-bench

SWE-bench: benchmark for evaluating LLMs on real-world GitHub issue resolution across 12 popular Python repositories.

99.15,762 stars · 957 forks · 9 mentions
Harnessclaudecursorcodexopencodegemini
evalscodebenchmarkagents
No one-command install · SourceDetails
SkillsPreview

trail-of-bits-security-skills

A very professional collection of over a dozen security-focused skills for code auditing and vulnerability detection. Includes skills for static analysis with CodeQL and Semgrep, variant analysis across codebases, fix verification, and differential code review.

99.06,939 stars · 597 forks
Harnessclaude
claude-codeagent-skills
No one-command install · SourceDetails
WorkflowsPreview

claude-code-ultimate-guide

A guide to Claude Code from beginner to power user, with templates for its features, guides on agentic workflows, quizzes and a cheatsheet. Check that it is current before relying on it.

99.05,869 stars · 770 forks
Harnessclaude
claude-codeworkflows-knowledge-guides
No one-command install · SourceDetails
EvalsPreview

giskard

Open-source LLM testing framework for detecting vulnerabilities (prompt injection, hallucinations, bias) via automated scan.

99.05,838 stars · 542 forks
Harnessclaudecursorcodexopencodegemini
evalssafetyvulnerabilityscan
No one-command install · SourceDetails
SkillsPreview

codebase-to-course

A Claude Code skill that turns any codebase into an interactive single-page HTML course for non-technical readers.

99.05,505 stars · 551 forks
Harnessclaude
claude-codeagent-skills
No one-command install · SourceDetails
WorkflowsExperimental

learn-agentic-ai

Learn Agentic AI using Dapr Agentic Cloud Ascent (DACA) Design Pattern: OpenAI Agents SDK, Memory, MCP, A2A, Knowledge Graphs, Rancher Desktop, and Kubernetes

99.04,351 stars · 1,008 forks
Harnessclaudecursorcodexopencodegemini
a2aagent-to-agenttutorials-learning-resources
No one-command install · SourceDetails
EvalsPreview

agenta

Open-source LLM developer platform with prompt playground, evaluation pipelines, and A/B testing for iterating on LLM apps.

98.94,670 stars · 661 forks
Harnessclaudecursorcodexopencodegemini
evalsplaygroundab-testingci
No one-command install · SourceDetails
EvalsPreview

openai-simple-evals

OpenAI's lightweight benchmark suite (MMLU, HumanEval, MATH, GPQA, MGSM) for fast model capability comparisons.

98.94,621 stars · 509 forks
Harnessclaudecursorcodexopencodegemini
evalsbenchmarkmmlusimple
No one-command install · SourceDetails
WorkflowsExperimental

ysymyth-react

[ICLR 2023] ReAct: Synergizing Reasoning and Acting in Language Models

Contributed by Sentinel

98.84,140 stars · 401 forks · 13 mentions
Harnessclaudecodexcursorgeminiopencode
No one-command install · SourceDetails
EvalsPreview

big-bench

Google's Beyond the Imitation Game benchmark: 200+ diverse tasks designed to probe capabilities beyond standard NLP benchmarks.

98.83,247 stars · 618 forks
Harnessclaudecursorcodexopencodegemini
evalsbenchmarkgoogleacademic
No one-command install · SourceDetails
EvalsPreview

trulens

Evaluation and tracking for LLM and RAG applications with a feedback-function API and experiment dashboard.

98.73,530 stars · 335 forks · 1 mention
Harnessclaudecursorcodexopencodegemini
evalsragtrackingdashboard
No one-command install · SourceDetails
EvalsExperimental

xlang-ai-osworld

[NeurIPS 2024] OSWorld: Benchmarking Multimodal Agents for Open-Ended Tasks in Real Computer Environments

Contributed by Sentinel

98.73,117 stars · 530 forks · 7 mentions
Harnessclaudecodexcursorgeminiopencode
No one-command install · SourceDetails
Browse · Armory