Armory
Source

Browse

Search and filter by type across the catalog

623 results in Infrastructure, Memory, Observability, Evals, CLAUDE.md / Rules · page 1 of 26

InfrastructurePreview

browser-use

Python library that makes web browsers accessible to AI agents; built on Playwright and LangChain. Supports multi-tab, vision + accessibility-tree hybrid mode, custom actions, and a self-correcting agent loop.

99.9111,989 stars · 12,312 forks · 11 mentions · passed install test
Harnessclaudecursorcodexopencodegemini
browserbrowser-use
No one-command install · SourceDetails
CLAUDE.md / RulesStable

karpathy-coding-discipline

Drop into CLAUDE.md/AGENTS.md as the first behavior norm a coding agent ingrains: think before coding, prefer the simplest solution, change only what you own, and execute toward the stated goal.

99.9209,417 stars · 21,315 forks
Harnessclaudecodexcursorgeminiopencode
disciplinecodingconstitutionbehavior-norm
Details
InfrastructurePreview

daytona

Secure and elastic sandboxes for running AI-generated code. The public repository stopped updating in June 2026, when development moved to a private codebase.

99.971,846 stars · 5,651 forks · 9 mentions · passed install test
Harnessclaudecursorcodexopencodegemini
infrastructuredev-environments
No one-command install · SourceDetails
InfrastructureExperimental

ggml-org-llama-cpp

LLM inference in C/C++

Contributed by Sentinel

99.9126,728 stars · 22,647 forks · 5 mentions
Harnessclaudecodexcursorgeminiopencode
No one-command install · SourceDetails
InfrastructureExperimental

vllm-project-vllm

A high-throughput and memory-efficient inference and serving engine for LLMs

Contributed by Sentinel

99.990,743 stars · 21,577 forks · 15 mentions
Harnessclaudecodexcursorgeminiopencode
No one-command install · SourceDetails
MemoryExperimental

thedotmack-claude-mem

Persistent Context Across Sessions for Every Agent – Captures everything your agent does during sessions, compresses it with AI, and injects relevant context back into future sessions. Works with Claude Code, OpenClaw, Codex, Gemini, Hermes, Copilot, OpenCode + More

Contributed by Sentinel

99.894,727 stars · 8,379 forks · 6 mentions
Harnessclaudecodexcursorgeminiopencode
cli
No one-command install · SourceDetails
InfrastructureStable

e2b-sandbox

Use as the default runtime when an agent must execute untrusted code or commands. Firecracker microVMs with ~150ms cold start give each run an isolated, disposable computer.

99.813,635 stars · 1,015 forks · 10 mentions · passed install test
Harnessclaudecodex
sandboxruntimefirecrackermicrovm
No one-command install · SourceDetails
InfrastructureExperimental

diegosouzapw-omniroute

Never stop coding. Free MIT AI gateway: one endpoint, 352 providers (150+ free), 1200+ models Kimi, Claude, GPT, Gemini, GLM, DeepSeek, MiniMax. Works with Claude Code, Codex, Cursor, OpenCode, Cline & Copilot. Quota-aware auto-fallback, RTK+Caveman compression saves 15-95% tokens, MCP/A2A, Desktop/PWA. Built by 550+ contributors

Contributed by Sentinel

99.860,005 stars · 8,342 forks · 3 mentions
Harnessclaudecodexcursorgeminiopencode
No one-command install · SourceDetails
MemoryExperimental

mempalace

Use when you want an agent memory system whose recall quality has actually been benchmarked rather than asserted.

99.858,897 stars · 7,548 forks · 2 mentions
Harnessclaudecodexcursorgeminiopencode
cp138-seedmemory
No one-command install · SourceDetails
InfrastructureExperimental

sgl-project-sglang

SGLang is a high-performance serving framework for large language models and multimodal models.

Contributed by Sentinel

99.733,202 stars · 8,472 forks · 4 mentions
Harnessclaudecodexcursorgeminiopencode
No one-command install · SourceDetails
ObservabilityPreview

sentry-llm-monitoring

Sentry's error and performance monitoring extended to LLM applications. It captures exceptions, latency, and AI token usage with OpenTelemetry integration.

99.744,709 stars · 4,832 forks · 15 mentions
Harnessclaudecursorcodexopencodegemini
observabilityerrorsapm
No one-command install · SourceDetails
InfrastructureStable

stripe-agent-toolkit

Use as the payments rail when an agent should earn or spend money in code (create customers, prices, payment links, and usage-based billing): the infrastructure behind an agent that funds its own compute.

99.71,785 stars · 329 forks · passed install test
Harnessclaudecodex
paymentsstripebillingfinancial-rails
No one-command install · SourceDetails
MemoryExperimental

microsoft-graphrag

A modular graph-based Retrieval-Augmented Generation (RAG) system

Contributed by Sentinel

99.635,783 stars · 3,754 forks · 3 mentions
Harnessclaudecodexcursorgeminiopencode
No one-command install · SourceDetails
ObservabilityPreview

mlflow-tracing

MLflow's LLM tracing module instruments model calls, agent steps, and tool invocations, storing them alongside experiment runs for reproducibility.

99.627,768 stars · 6,246 forks
Harnessclaudecursorcodexopencodegemini
observabilitytracingexperiment-tracking
No one-command install · SourceDetails
ObservabilityPreview

langfuse

Open-source LLM engineering platform with traces, evals, prompt management, and datasets for debugging and improving LLM applications.

99.634,067 stars · 3,678 forks · 10 mentions
Harnessclaudecursorcodexopencodegemini
observabilitytracingevals
No one-command install · SourceDetails
MemoryExperimental

volcengine-openviking

Use when an agent's memory, retrieved knowledge and learned skills should live in one store that reorganises itself.

99.635,893 stars · 2,739 forks · 2 mentions
Harnessclaudecodexcursorgeminiopencode
cp138-seedmemory
No one-command install · SourceDetails
MemoryExperimental

garrytan-gbrain

Garry's Opinionated OpenClaw/Hermes Agent Brain

Contributed by Sentinel

99.629,460 stars · 4,394 forks · 16 mentions
Harnessclaudecodexcursorgeminiopencode
No one-command install · SourceDetails
MemoryExperimental

getzep-graphiti

Build Real-Time Knowledge Graphs for AI Agents

Contributed by Sentinel

99.630,509 stars · 3,098 forks · 1 mention
Harnessclaudecodexcursorgeminiopencode
memory
No one-command install · SourceDetails
MemoryExperimental

topoteretes-cognee

Use when an agent needs long-term memory backed by a knowledge graph you can host yourself.

99.630,556 stars · 3,010 forks · 4 mentions
Harnessclaudecodexcursorgeminiopencode
cp138-seedmemory
No one-command install · SourceDetails
MemoryExperimental

supermemoryai-supermemory

Use when memory has to be fast, run locally, and be reachable from an app as well as an agent.

99.629,256 stars · 2,557 forks
Harnessclaudecodexcursorgeminiopencode
cp138-seedmemory
No one-command install · SourceDetails
MemoryExperimental

tencentcloud-tencentdb-agent-memory

TencentDB Agent Memory is a team-level memory hub for AI Agents — turning conversations, docs, and code into four reusable memory assets (Chat Memory, Skill, LLM-Wiki, Code-Graph) that are governed, shared, and equipped across agents and frameworks.

Contributed by Sentinel

99.525,635 stars · 2,398 forks · 3 mentions
Harnessclaudecodexcursorgeminiopencode
No one-command install · SourceDetails
MemoryExperimental

letta-ai-letta

Platform for stateful agents: AI with advanced memory that can learn and self-improve over time.

Contributed by Sentinel

99.524,552 stars · 2,609 forks · 19 mentions
Harnessclaudecodexcursorgeminiopencode
memory
No one-command install · SourceDetails
CLAUDE.md / RulesExperimental

humanlayer-12-factor-agents

What are the principles we can use to build LLM-powered software that is actually good enough to put in the hands of production customers?

Contributed by Sentinel

99.525,642 stars · 1,951 forks · 4 mentions
Harnessclaudecodexcursorgeminiopencode
No one-command install · SourceDetails
EvalsPreview

promptfoo

CLI and library for testing, comparing, and red-teaming LLM prompts and agents with assertions and CI integration.

99.524,737 stars · 2,255 forks
Harnessclaudecursorcodexopencodegemini
evalsred-teamingcicli
No one-command install · SourceDetails
Browse · Armory