Armory
Source

Browse

Search and filter by type across the catalog

1,166 results in Skills, Observability · page 3 of 49

SkillsPreview

claude-code-agents

Comprehensive E2E development workflow with helpful Claude Code subagent prompts for solo devs. Run multiple auditors in parallel, automate fix cycles with micro-checkpoint protocols, and do browser-based QA. Includes strict protocols to prevent AI going rogue.

91.9147 stars · 15 forks
Harnessclaude
skill
No one-command install · SourceDetails
SkillsPreview

book-factory

A comprehensive pipeline of Skills that replicates traditional publishing infrastructure for nonfiction book creation using specialized Claude skills.

91.6108 stars · 20 forks
Harnessclaude
skill
No one-command install · SourceDetails
ObservabilityPreview

claudia-statusline

High-performance Rust-based statusline for Claude Code with persistent stats tracking, progress bars, and optional cloud sync. Features SQLite-first persistence, git integration, context progress bars, burn rate calculation, XDG-compliant with theme support (dark/light, NO_COLOR).

81.236 stars · 5 forks
Harnessclaude
claude-codestatus-lines
No one-command install · SourceDetails
ObservabilityPreview

honeycomb

Honeycomb's OpenTelemetry-native observability platform: a high-cardinality event store ideal for tracing LLM pipelines and debugging slow agent traces.

75.116 stars · 6 forks
Harnessclaudecursorcodexopencodegemini
observabilityopentelemetrytracing
No one-command install · SourceDetails
ObservabilityPreview

baserun

Baserun captures LLM traces via a lightweight decorator-based SDK and provides a dashboard for debugging prompt chains, testing variants, and measuring quality.

74.416 stars · 5 forks
Harnessclaudecursorcodexopencodegemini
observabilitytracingtesting
No one-command install · SourceDetails
SkillsPreview

claude-mountaineering-skills

Claude Code skill that automates mountain route research for North American peaks. Aggregates data from 10+ mountaineering sources like Mountaineers.org, PeakBagger.com and SummitPost.com to generate detailed route beta reports with weather, avalanche conditions, and trip reports.

73.333 stars · 1 fork
Harnessclaude
skill
No one-command install · SourceDetails
SkillsPreview

read-only-postgres

Read-only PostgreSQL query skill for Claude Code. Executes SELECT/SHOW/EXPLAIN/WITH queries across configured databases with strict validation, timeouts, and row limits. Supports multiple connections with descriptions for database selection.

65.614 stars · 1 fork
Harnessclaude
claude-codeagent-skills
No one-command install · SourceDetails
SkillsPreview

2d-games

2D game development principles. Sprites, tilemaps, physics, camera.

UnrankedNo signals yet
Harnessclaude
2d-gamesskills
Details
SkillsPreview

3d-games

3D game development principles. Rendering, shaders, physics, cameras.

UnrankedNo signals yet
Harnessclaude
3d-gamesskills
Details
SkillsPreview

3d-web-experience

Expert in building 3D experiences for the web - Three.js, React Three Fiber, Spline, WebGL, and interactive 3D scenes. Covers product configurators, 3D portfolios, immersive websites, and bringing depth to web experiences. Use when: 3D website, three.js, WebGL, react three fiber, 3D experience.

UnrankedNo signals yet
Harnessclaude
3d-web-experienceskills
Details
SkillsPreview

ab-test-setup

When the user wants to plan, design, or implement an A/B test or experiment. Also use when the user mentions "A/B test," "split test," "experiment," "test this change," "variant copy," "multivariate test," or "hypothesis." For tracking implementation, see analytics-tracking.

UnrankedNo signals yet
Harnessclaude
ab-test-setupskills
Details
SkillsPreview

academic-cv-builder

Format CVs for academic positions, including publications, grants, teaching, and research experience. Use when the user mentions academic CV, faculty job, tenure-track, postdoc, PhD, research position, or needs to list publications, grants, conferences, or teaching.

UnrankedNo signals yet
Harnessclaude
academic-cv-builderskills
Details
SkillsPreview

accessibility

Design, implement, and audit inclusive digital products using WCAG 2.2 Level AA standards. Use this skill to generate semantic ARIA for Web and accessibility traits for Web and Native platforms (iOS/Android).

UnrankedNo signals yet
Harnessclaudecodexcursorgeminiopencode
skill
Details
SkillsPreview

accessibility-2

Audit and improve web accessibility following WCAG 2.1 guidelines. Use when asked to "improve accessibility", "a11y audit", "WCAG compliance", "screen reader support", "keyboard navigation", or "make accessible".

UnrankedNo signals yet
Harnessclaude
accessibilityskills
Details
SkillsPreview

accessibility-auditor

Web accessibility specialist for WCAG compliance, ARIA implementation, and inclusive design. Use when auditing websites for accessibility issues, implementing WCAG 2.1 AA/AAA standards, testing with screen readers, or ensuring ADA compliance. Expert in semantic HTML, keyboard navigation, and assistive technology compatibility.

UnrankedNo signals yet
Harnessclaude
accessibility-auditorskills
Details
SkillsPreview

active-directory-attacks

This skill should be used when the user asks to "attack Active Directory", "exploit AD", "Kerberoasting", "DCSync", "pass-the-hash", "BloodHound enumeration", "Golden Ticket", "Silver Ticket", "AS-REP roasting", "NTLM relay", or needs guidance on Windows domain penetration testing.

UnrankedNo signals yet
Harnessclaude
active-directory-attacksskills
Details
SkillsPreview

adaptyv

Cloud laboratory platform for automated protein testing and validation. Use when designing proteins and needing experimental validation including binding assays, expression testing, thermostability measurements, enzyme activity assays, or protein sequence optimization. Also use for submitting experiments via API, tracking experiment status, downloading results, optimizing protein sequences for better expression using computational tools (NetSolP, SoluProt, SolubleMPNN, ESM), or managing protein design workflows with wet-lab validation.

UnrankedNo signals yet
Harnessclaude
adaptyvskills
Details
SkillsPreview

address-github-comments

Use when you need to address review or issue comments on an open GitHub Pull Request using the gh CLI.

UnrankedNo signals yet
Harnessclaude
address-github-commentsskills
Details
SkillsPreview

aeon

This skill should be used for time series machine learning tasks including classification, regression, clustering, forecasting, anomaly detection, segmentation, and similarity search. Use when working with temporal data, sequential patterns, or time-indexed observations requiring specialized algorithms beyond standard ML approaches. Particularly suited for univariate and multivariate time series analysis with scikit-learn compatible APIs.

UnrankedNo signals yet
Harnessclaude
aeonskills
Details
SkillsPreview

agent-architecture-audit

Full-stack diagnostic for agent and LLM applications. Audits the 12-layer agent stack for wrapper regression, memory pollution, tool discipline failures, hidden repair loops, and rendering corruption. Produces severity-ranked findings with code-first fixes. Essential for developers building agent…

UnrankedNo signals yet
Harnessclaudecodexcursorgeminiopencode
skill
Details
SkillsPreview

agent-development

This skill should be used when the user asks to "create an agent", "add an agent", "write a subagent", "agent frontmatter", "when to use description", "agent examples", "agent tools", "agent colors", "autonomous agent", or needs guidance on agent structure, system prompts, triggering conditions, or agent development best practices for Claude Code plugins.

UnrankedNo signals yet
Harnessclaude
agent-developmentskills
Details
SkillsPreview

agent-eval

Head-to-head comparison of coding agents (Claude Code, Aider, Codex, etc.) on custom tasks with pass rate, cost, time, and consistency metrics

UnrankedNo signals yet
Harnessclaudecodexcursorgeminiopencode
agentsevalbenchmark
Details
SkillsPreview

agent-evaluation

Testing and benchmarking LLM agents including behavioral testing, capability assessment, reliability metrics, and production monitoring—where even top agents achieve less than 50% on real-world benchmarks Use when: agent testing, agent evaluation, benchmark agents, agent reliability, test agent.

UnrankedNo signals yet
Harnessclaude
agent-evaluationskills
Details
SkillsPreview

agent-harness-construction

Design and optimize AI agent action spaces, tool definitions, and observation formatting for higher completion rates.

UnrankedNo signals yet
Harnessclaudecodexcursorgeminiopencode
harnessagentstools
Details
Browse · Armory