Catalogue
Submit a toolHarnesses, frameworks, tools, apps, and platforms for agent builders, each scored on real GitHub credibility and each with a credibility-gated forum.
87 tools · showing 49–72
ToolOpenWolf
MemoryMonitoring
cytostack
OpenWolf treats Claude Code's token burn as a caching problem … six hook scripts intercept every read/write, maintaining
clilocal-firsttoken-optimizationobservability
ToolAIProxy
CodingMonitoring
labring
AI gateway with OpenAI-, Anthropic-, and Gemini-compatible routing, multi-tenant management, monitoring, and a plugin system
gatewaymulti-tenantopenai-compatiblemonitoring-dashboard
PlatformBraintrust
QAMonitoring
braintrustdata
Eval, logging and prompt-playground platform for LLM and agent applications.
evaluationobservabilityprompt
ToolAgentacct
Monitoring
mikehasa
Your coding agent says a task is done. Before you trust that, agentacct checks it, readin...
token-optimizationcoding-agentlocal-firstgui
AlphaClaw
DeploymentMonitoring
chrysb
Web dashboard around OpenClaw with a setup wizard, self-healing watchdog, Git-backed rollback, and multi-agent management
watchdogself-healingrollbackmulti-agent-management
ToolHomeRail
VoiceCodingMonitoring
xiaotianfotos
Voice-first local agent orchestration runtime for auditable DAG workflows.
local-firstdagorchestrationhome-lab
HarnessSandBase Harness
SecurityMonitoring
sandbaseai
Open-source CMA-compatible agent runtime for any model, with MCP tools, sandboxed sessions, audit, replay, and a local console. Includes a n
sandboxlocal-firstgovernancesession-replay
Vigolium
MonitoringSecurity
vigolium
Vigolium - High-fidelity vulnerability scanner fusing agentic AI with native speed, modularity, and precision
vulnerability-scanneroast
ToolGini Agent
MemoryInterfaceMonitoring
Open-Curiosity
The agent that remembers and learns.
productivityassistantlocal-firstskillhuman-in-the-loop
ToolRaindrop Workshop
QAMonitoringCoding
raindrop-ai
Open-source tool that lets a coding agent write and run agent evals locally
observabilityevaluationdebugginglocal-first
ToolLLM Wiki
MemoryMonitoring
Pratiyush
LLM-powered knowledge base from your Claude Code, Codex CLI, Copilot, Cursor & Gemini sessions. Karpathy's LLM Wiki pattern — implemented an
knowledge-basesession-managementguistatic-site
ToolPandaProbe
CodingMonitoring
chirpz-ai
Open-source platform for tracing, evaluating, and monitoring AI agents, with integrations for LangGraph, CrewAI, and agent SDKs
observabilityevaluationlangchaincrewai
ToolCascadeFlow
InferenceMonitoring
lemony-ai
Cascading runtime for AI agents. Optimize cost, latency, quality, and policy decisions inside the agent loop.
model-cascadingtoken-optimizationgatewaylangchain
Helicone
InferenceMonitoring
Acquired
Helicone
AI gateway to 100+ models with routing and fallbacks, plus request logging, tracing, and cost analytics
gatewayfallbackscost-analyticsobservability
ToolNumbat
SecurityMonitoring
perplexityai
Visibility into AI agent activity on endpoints, with on-device detection, optional pre-action blocking, and forensic reconstruction.
endpoint-detectionforensicsgovernancetelemetry
ToolAdrian
SecurityMonitoring
secureagentics
Say your agent starts resetting passwords it shouldn't. Logging it after the fact is too late, and no prompt-injection c
prompt-injectionruntime-securitythreat-detectionpolicy-drift
ToolScorable
QAMonitoring
root-signals
Evaluation platform (formerly Root Signals) exposing judges and evaluators to agents via SDK and MCP for in-loop quality scoring.
evaluationsdk
ToolTokenTelemetry
MonitoringCoding
VasiHemanth
Token telemetry dashboard for AI autonomous and coding agents — tracks tokens, sessions, tool calls & reasoning across Hermes agent, Claude
local-firstobservabilitytelemetrytoken-optimization
ToolTracely
MonitoringQADeployment
Jwuthri
You fix the agent bug and write the regression test. Then you try to reproduce the run: t...
observabilityci-cdevaluationregression-testing
ToolAgent Flow
Monitoring
patoles
Real-time visualization of Claude Code agent orchestration — see your agents think, branch, and coordinate as they work.
agent-visualizationobservabilitydebuggingide
ToolAI Observer
Monitoring
tobilg
Unified local observability for AI coding assistants
telemetryduckdbtoken-optimizationgui
ToolITOps Agent Platform
DeploymentCodingMonitoring
qinshihu
China's 1st enterprise multi-agent IT ops platform. LLM-powered auto-remediation for Zabbix/Prometheus. Docker deploy.
multi-agentincident-managementdebuggingdocker
ToolAgentOps
CodingMonitoring
AgentOps-AI
Python SDK for AI agent monitoring, LLM cost tracking, benchmarking, and more. Integrates with most LLMs and agent frameworks including CrewAI, Agno, OpenAI Agents SDK, Langchain, Autogen, AG2, and CamelAI.
observabilitytoken-optimizationcrewailangchain
ToolEvidently
QAMonitoring
evidentlyai
Open-source Python framework to evaluate, test, and monitor ML and LLM systems from experiments to production
data-driftobservabilityevaluationllmops