Armory
Source

Browse

Search and filter by type across the catalog

1,310 results in Identity, Infrastructure, Workflows, CLAUDE.md / Rules, Evals · page 1 of 55

WorkflowsExperimental

n8n-io-n8n

Fair-code workflow automation platform with native AI capabilities. Combine visual building with custom code, self-host or cloud, 400+ integrations.

Contributed by Sentinel

99.9206,033 stars · 60,896 forks · 26 mentions
Harnessclaudecodexcursorgeminiopencode
No one-command install · SourceDetails
InfrastructurePreview

browser-use

Python library that makes web browsers accessible to AI agents; built on Playwright and LangChain. Supports multi-tab, vision + accessibility-tree hybrid mode, custom actions, and a self-correcting agent loop.

99.9111,989 stars · 12,312 forks · 11 mentions · passed install test
Harnessclaudecursorcodexopencodegemini
browserbrowser-use
No one-command install · SourceDetails
CLAUDE.md / RulesStable

karpathy-coding-discipline

Drop into CLAUDE.md/AGENTS.md as the first behavior norm a coding agent ingrains: think before coding, prefer the simplest solution, change only what you own, and execute toward the stated goal.

99.9209,417 stars · 21,315 forks
Harnessclaudecodexcursorgeminiopencode
disciplinecodingconstitutionbehavior-norm
Details
InfrastructurePreview

daytona

Secure and elastic sandboxes for running AI-generated code. The public repository stopped updating in June 2026, when development moved to a private codebase.

99.971,846 stars · 5,651 forks · 9 mentions · passed install test
Harnessclaudecursorcodexopencodegemini
infrastructuredev-environments
No one-command install · SourceDetails
InfrastructureExperimental

ggml-org-llama-cpp

LLM inference in C/C++

Contributed by Sentinel

99.9126,728 stars · 22,647 forks · 5 mentions
Harnessclaudecodexcursorgeminiopencode
No one-command install · SourceDetails
InfrastructureExperimental

vllm-project-vllm

A high-throughput and memory-efficient inference and serving engine for LLMs

Contributed by Sentinel

99.990,743 stars · 21,577 forks · 15 mentions
Harnessclaudecodexcursorgeminiopencode
No one-command install · SourceDetails
WorkflowsExperimental

karpathy-autoresearch

AI agents running research on single-GPU nanochat training automatically

Contributed by Sentinel

99.995,090 stars · 13,383 forks · 7 mentions
Harnessclaudecodexcursorgeminiopencode
No one-command install · SourceDetails
InfrastructureStable

e2b-sandbox

Use as the default runtime when an agent must execute untrusted code or commands. Firecracker microVMs with ~150ms cold start give each run an isolated, disposable computer.

99.813,635 stars · 1,015 forks · 10 mentions · passed install test
Harnessclaudecodex
sandboxruntimefirecrackermicrovm
No one-command install · SourceDetails
WorkflowsPreview

learn-claude-code

An analysis of how coding agents like Claude Code are designed, which breaks an agent into its basic parts and rebuilds it with minimal code: a rudimentary agent with skills, sub-agents and a to-do list in a few hundred lines of Python.

99.875,849 stars · 12,224 forks
Harnessclaude
claude-codeworkflows-knowledge-guides
No one-command install · SourceDetails
InfrastructureExperimental

diegosouzapw-omniroute

Never stop coding. Free MIT AI gateway: one endpoint, 352 providers (150+ free), 1200+ models Kimi, Claude, GPT, Gemini, GLM, DeepSeek, MiniMax. Works with Claude Code, Codex, Cursor, OpenCode, Cline & Copilot. Quota-aware auto-fallback, RTK+Caveman compression saves 15-95% tokens, MCP/A2A, Desktop/PWA. Built by 550+ contributors

Contributed by Sentinel

99.860,005 stars · 8,342 forks · 3 mentions
Harnessclaudecodexcursorgeminiopencode
No one-command install · SourceDetails
InfrastructureExperimental

sgl-project-sglang

SGLang is a high-performance serving framework for large language models and multimodal models.

Contributed by Sentinel

99.733,202 stars · 8,472 forks · 4 mentions
Harnessclaudecodexcursorgeminiopencode
No one-command install · SourceDetails
InfrastructureStable

stripe-agent-toolkit

Use as the payments rail when an agent should earn or spend money in code (create customers, prices, payment links, and usage-based billing): the infrastructure behind an agent that funds its own compute.

99.71,785 stars · 329 forks · passed install test
Harnessclaudecodex
paymentsstripebillingfinancial-rails
No one-command install · SourceDetails
WorkflowsExperimental

openai-symphony

Symphony turns project work into isolated, autonomous implementation runs, allowing teams to manage work instead of supervising coding agents.

Contributed by Sentinel

99.526,991 stars · 2,771 forks · 4 mentions
Harnessclaudecodexcursorgeminiopencode
No one-command install · SourceDetails
WorkflowsExperimental

a2a-protocol-github-repository

Google's official repository for A2A protocol

99.525,941 stars · 2,633 forks
Harnessclaudecursorcodexopencodegemini
a2aagent-to-agentofficial-resources
No one-command install · SourceDetails
CLAUDE.md / RulesExperimental

humanlayer-12-factor-agents

What are the principles we can use to build LLM-powered software that is actually good enough to put in the hands of production customers?

Contributed by Sentinel

99.525,642 stars · 1,951 forks · 4 mentions
Harnessclaudecodexcursorgeminiopencode
No one-command install · SourceDetails
EvalsPreview

promptfoo

CLI and library for testing, comparing, and red-teaming LLM prompts and agents with assertions and CI integration.

99.524,737 stars · 2,255 forks
Harnessclaudecursorcodexopencodegemini
evalsred-teamingcicli
No one-command install · SourceDetails
InfrastructurePreview

skyvern

Open-source agent platform that automates browser-based workflows using LLMs and computer vision. It identifies interactive elements via screenshots, handles CAPTCHAs, and supports complex multi-step form flows.

99.522,907 stars · 2,152 forks
Harnessclaudecursorcodexopencodegemini
browserskyvern
No one-command install · SourceDetails
EvalsPreview

openai-evals

OpenAI's official framework for evaluating LLMs and LLM-powered systems, with a registry of community eval sets.

99.519,509 stars · 3,093 forks
Harnessclaudecursorcodexopencodegemini
evalsregistrybenchmark
No one-command install · SourceDetails
EvalsPreview

lm-evaluation-harness

EleutherAI's unified framework for evaluating language models on hundreds of academic benchmarks.

99.513,860 stars · 3,533 forks
Harnessclaudecursorcodexopencodegemini
evalsacademicbenchmarkharness
No one-command install · SourceDetails
InfrastructureStable

browserbase-bb

Use when an agent must operate the live web (navigate, act, and extract on real pages) via a cloud browser driven by act/extract/observe primitives, with a local-Chromium escape hatch using the same code.

99.424,125 stars · 1,664 forks · 3 mentions
Harnessclaudecodex
browserweb-automationstagehandbrowserbase
No one-command install · SourceDetails
WorkflowsExperimental

anthropic-quickstarts

Offers comprehensive development guides for three distinct AI-powered demo projects with standardized workflows, strict code style guidelines, and containerization instructions.

99.417,588 stars · 3,031 forks
Harnessclaude
awesome-claude-codeofficial-documentation
No one-command install · SourceDetails
InfrastructurePreview

browser-use-webui

Gradio web UI on top of the browser-use framework. It lets users run AI browser agents interactively, configure LLM providers, watch live recordings, and replay task sessions without writing Python.

99.416,310 stars · 2,721 forks
Harnessclaudecursorcodexopencodegemini
browserbrowser-use
No one-command install · SourceDetails
InfrastructurePreview

cua-computer-use-agent

trycua/cua open-source computer-use agent framework: Apple Silicon-native, runs lightweight macOS/Linux VMs with sub-second cold starts; provides a unified Python interface for screen capture, click, and type actions.

99.422,092 stars · 1,519 forks
Harnessclaudecursorcodexopencodegemini
browsercomputer-use
No one-command install · SourceDetails
EvalsPreview

deepeval

Open-source LLM evaluation framework with 14+ metrics (hallucination, faithfulness, answer relevancy) and CI support.

99.418,041 stars · 1,890 forks
Harnessclaudecursorcodexopencodegemini
evalsmetricsragci
No one-command install · SourceDetails
Browse · Armory