1,292 results in CLAUDE.md / Rules, Memory, Evals, Observability, Sub-Agents · page 1 of 54
Harness Claude Code Cursor Codex Gemini OpenCode karpathy-coding-discipline Drop into CLAUDE.md/AGENTS.md as the first behavior norm a coding agent ingrains: think before coding, prefer the simplest solution, change only what you own, and execute toward the stated goal.
99.9 209,417 stars · 21,315 forks
Harness claude codex cursor gemini opencode
discipline coding constitution behavior-norm
thedotmack-claude-mem Persistent Context Across Sessions for Every Agent – Captures everything your agent does during sessions, compresses it with AI, and injects relevant context back into future sessions. Works with Claude Code, OpenClaw, Codex, Gemini, Hermes, Copilot, OpenCode + More
Contributed by Sentinel
99.8 94,727 stars · 8,379 forks · 6 mentions
Harness claude codex cursor gemini opencode
cli
mempalace Use when you want an agent memory system whose recall quality has actually been benchmarked rather than asserted.
99.8 58,897 stars · 7,548 forks · 2 mentions
Harness claude codex cursor gemini opencode
cp138-seed memory
sentry-llm-monitoring Sentry's error and performance monitoring extended to LLM applications. It captures exceptions, latency, and AI token usage with OpenTelemetry integration.
99.7 44,709 stars · 4,832 forks · 15 mentions
Harness claude cursor codex opencode gemini
observability errors apm
wshobson-agents Multi-harness agentic plugin marketplace for Claude Code, Codex, Cursor, OpenCode, GitHub Copilot, and Google Antigravity
Contributed by Sentinel
99.7 39,338 stars · 4,195 forks
Harness claude codex cursor gemini opencode
subagent
microsoft-graphrag A modular graph-based Retrieval-Augmented Generation (RAG) system
Contributed by Sentinel
99.6 35,783 stars · 3,754 forks · 3 mentions
Harness claude codex cursor gemini opencode
mlflow-tracing MLflow's LLM tracing module instruments model calls, agent steps, and tool invocations, storing them alongside experiment runs for reproducibility.
99.6 27,768 stars · 6,246 forks
Harness claude cursor codex opencode gemini
observability tracing experiment-tracking
langfuse Open-source LLM engineering platform with traces, evals, prompt management, and datasets for debugging and improving LLM applications.
99.6 34,067 stars · 3,678 forks · 10 mentions
Harness claude cursor codex opencode gemini
observability tracing evals
volcengine-openviking Use when an agent's memory, retrieved knowledge and learned skills should live in one store that reorganises itself.
99.6 35,893 stars · 2,739 forks · 2 mentions
Harness claude codex cursor gemini opencode
cp138-seed memory
garrytan-gbrain Garry's Opinionated OpenClaw/Hermes Agent Brain
Contributed by Sentinel
99.6 29,460 stars · 4,394 forks · 16 mentions
Harness claude codex cursor gemini opencode
getzep-graphiti Build Real-Time Knowledge Graphs for AI Agents
Contributed by Sentinel
99.6 30,509 stars · 3,098 forks · 1 mention
Harness claude codex cursor gemini opencode
memory
topoteretes-cognee Use when an agent needs long-term memory backed by a knowledge graph you can host yourself.
99.6 30,556 stars · 3,010 forks · 4 mentions
Harness claude codex cursor gemini opencode
cp138-seed memory
supermemoryai-supermemory Use when memory has to be fast, run locally, and be reachable from an app as well as an agent.
99.6 29,256 stars · 2,557 forks
Harness claude codex cursor gemini opencode
cp138-seed memory
voltagent-awesome-claude-code-subagents A collection of 100+ specialized Claude Code subagents covering a wide range of development use cases
Contributed by Sentinel
99.5 24,793 stars · 2,870 forks
Harness claude codex cursor gemini opencode
subagent
tencentcloud-tencentdb-agent-memory TencentDB Agent Memory is a team-level memory hub for AI Agents — turning conversations, docs, and code into four reusable memory assets (Chat Memory, Skill, LLM-Wiki, Code-Graph) that are governed, shared, and equipped across agents and frameworks.
Contributed by Sentinel
99.5 25,635 stars · 2,398 forks · 3 mentions
Harness claude codex cursor gemini opencode
letta-ai-letta Platform for stateful agents: AI with advanced memory that can learn and self-improve over time.
Contributed by Sentinel
99.5 24,552 stars · 2,609 forks · 19 mentions
Harness claude codex cursor gemini opencode
memory
CLAUDE.md / Rules Experimental humanlayer-12-factor-agents What are the principles we can use to build LLM-powered software that is actually good enough to put in the hands of production customers?
Contributed by Sentinel
99.5 25,642 stars · 1,951 forks · 4 mentions
Harness claude codex cursor gemini opencode
promptfoo CLI and library for testing, comparing, and red-teaming LLM prompts and agents with assertions and CI integration.
99.5 24,737 stars · 2,255 forks
Harness claude cursor codex opencode gemini
evals red-teaming ci cli
claude-hud A status line for Claude Code that shows context usage, tools, agents, to-dos and more. Highly configurable, and maintained when it was listed.
99.5 27,778 stars · 1,286 forks
Harness claude
statusline observability
openai-evals OpenAI's official framework for evaluating LLMs and LLM-powered systems, with a registry of community eval sets.
99.5 19,509 stars · 3,093 forks
Harness claude cursor codex opencode gemini
evals registry benchmark
lm-evaluation-harness EleutherAI's unified framework for evaluating language models on hundreds of academic benchmarks.
99.5 13,860 stars · 3,533 forks
Harness claude cursor codex opencode gemini
evals academic benchmark harness
gibsonai-memori Memori is agent-native memory infrastructure. A LLM-agnostic layer that turns agent execution and conversation into structured, persistent state for production systems. Built for enterprise, Memori works with the data infrastructure you already run, no rip-and-replace, and deploys across managed cloud, single-tenant cloud, VPC, and on-premises.
Contributed by Sentinel
99.4 16,313 stars · 3,283 forks
Harness claude codex cursor gemini opencode
memory
comet-opik Opik by Comet is an open-source LLM evaluation and tracing platform: log traces, run automated evals, create datasets, and track prompt improvements over time.
99.4 22,249 stars · 1,827 forks · 1 mention
Harness claude cursor codex opencode gemini
observability evals tracing
opentelemetry-genai OpenTelemetry semantic conventions and instrumentation for GenAI/LLM spans, traces, and metrics via the GenAI semconv working group.
99.4 4,897 stars · 3,851 forks
Harness claude cursor codex opencode gemini
observability opentelemetry tracing
Page 1 of 54 Next
Browse · Armory