This is an early release preview. You may encounter bugs.

Catalogue

Submit a tool

Harnesses, frameworks, tools, apps, and platforms for agent builders, each scored on real GitHub credibility and each with a credibility-gated forum.

87 tools · showing 49–72

monitoring ×
Tool

OpenWolf

MemoryMonitoring

cytostack

OpenWolf treats Claude Code's token burn as a caching problem … six hook scripts intercept every read/write, maintaining

clilocal-firsttoken-optimizationobservability

A
Tool

AIProxy

CodingMonitoring

labring

AI gateway with OpenAI-, Anthropic-, and Gemini-compatible routing, multi-tenant management, monitoring, and a plugin system

gatewaymulti-tenantopenai-compatiblemonitoring-dashboard

B
Platform

Braintrust

QAMonitoring

braintrustdata

Eval, logging and prompt-playground platform for LLM and agent applications.

evaluationobservabilityprompt

B
Tool

Agentacct

Monitoring

mikehasa

Your coding agent says a task is done. Before you trust that, agentacct checks it, readin...

token-optimizationcoding-agentlocal-firstgui

B
Harness

AlphaClaw

DeploymentMonitoring

chrysb

Web dashboard around OpenClaw with a setup wizard, self-healing watchdog, Git-backed rollback, and multi-agent management

watchdogself-healingrollbackmulti-agent-management

B
Tool

HomeRail

VoiceCodingMonitoring

xiaotianfotos

Voice-first local agent orchestration runtime for auditable DAG workflows.

local-firstdagorchestrationhome-lab

B
Harness

SandBase Harness

SecurityMonitoring

sandbaseai

Open-source CMA-compatible agent runtime for any model, with MCP tools, sandboxed sessions, audit, replay, and a local console. Includes a n

sandboxlocal-firstgovernancesession-replay

B
Tool

Vigolium

MonitoringSecurity

vigolium

Vigolium - High-fidelity vulnerability scanner fusing agentic AI with native speed, modularity, and precision

vulnerability-scanneroast

B
Tool

Gini Agent

MemoryInterfaceMonitoring

Open-Curiosity

The agent that remembers and learns.

productivityassistantlocal-firstskillhuman-in-the-loop

B
Tool

Raindrop Workshop

QAMonitoringCoding

raindrop-ai

Open-source tool that lets a coding agent write and run agent evals locally

observabilityevaluationdebugginglocal-first

B
Tool

LLM Wiki

MemoryMonitoring

Pratiyush

LLM-powered knowledge base from your Claude Code, Codex CLI, Copilot, Cursor & Gemini sessions. Karpathy's LLM Wiki pattern — implemented an

knowledge-basesession-managementguistatic-site

B
Tool

PandaProbe

CodingMonitoring

chirpz-ai

Open-source platform for tracing, evaluating, and monitoring AI agents, with integrations for LangGraph, CrewAI, and agent SDKs

observabilityevaluationlangchaincrewai

B
Tool

CascadeFlow

InferenceMonitoring

lemony-ai

Cascading runtime for AI agents. Optimize cost, latency, quality, and policy decisions inside the agent loop.

model-cascadingtoken-optimizationgatewaylangchain

B
Platform

Helicone

InferenceMonitoring

Acquired

Helicone

AI gateway to 100+ models with routing and fallbacks, plus request logging, tracing, and cost analytics

gatewayfallbackscost-analyticsobservability

B
Tool

Numbat

SecurityMonitoring

perplexityai

Visibility into AI agent activity on endpoints, with on-device detection, optional pre-action blocking, and forensic reconstruction.

endpoint-detectionforensicsgovernancetelemetry

B
Tool

Adrian

SecurityMonitoring

secureagentics

Say your agent starts resetting passwords it shouldn't. Logging it after the fact is too late, and no prompt-injection c

prompt-injectionruntime-securitythreat-detectionpolicy-drift

B
Tool

Scorable

QAMonitoring

root-signals

Evaluation platform (formerly Root Signals) exposing judges and evaluators to agents via SDK and MCP for in-loop quality scoring.

evaluationsdk

B
Tool

TokenTelemetry

MonitoringCoding

VasiHemanth

Token telemetry dashboard for AI autonomous and coding agents — tracks tokens, sessions, tool calls & reasoning across Hermes agent, Claude

local-firstobservabilitytelemetrytoken-optimization

B
Tool

Tracely

MonitoringQADeployment

Jwuthri

You fix the agent bug and write the regression test. Then you try to reproduce the run: t...

observabilityci-cdevaluationregression-testing

B
Tool

Agent Flow

Monitoring

patoles

Real-time visualization of Claude Code agent orchestration — see your agents think, branch, and coordinate as they work.

agent-visualizationobservabilitydebuggingide

B
Tool

AI Observer

Monitoring

tobilg

Unified local observability for AI coding assistants

telemetryduckdbtoken-optimizationgui

B
Tool

ITOps Agent Platform

DeploymentCodingMonitoring

qinshihu

China's 1st enterprise multi-agent IT ops platform. LLM-powered auto-remediation for Zabbix/Prometheus. Docker deploy.

multi-agentincident-managementdebuggingdocker

B
Tool

AgentOps

CodingMonitoring

AgentOps-AI

Python SDK for AI agent monitoring, LLM cost tracking, benchmarking, and more. Integrates with most LLMs and agent frameworks including CrewAI, Agno, OpenAI Agents SDK, Langchain, Autogen, AG2, and CamelAI.

observabilitytoken-optimizationcrewailangchain

B
Tool

Evidently

QAMonitoring

evidentlyai

Open-source Python framework to evaluate, test, and monitor ML and LLM systems from experiments to production

data-driftobservabilityevaluationllmops

B