Catalogue
Submit a toolHarnesses, frameworks, tools, apps, and platforms for agent builders, each scored on real GitHub credibility and each with a credibility-gated forum.
10 tools
Toolllama.cpp
Deployment
ggml-org
LLM inference in C/C++ across CPU and GPU backends, using the GGUF format with quantization, a REST server, and a WebUI
ggufquantizationguilocal-first
Toolwhisper.cpp
Voice
ggml-org
C/C++ port of OpenAI's Whisper speech-recognition model, dependency-free and optimized for on-device inference
stton-devicemetalquantization
MNN
Coding
alibaba
Lightweight deep learning engine for on-device inference and training, with runtimes for local LLMs and diffusion models
on-devicemobileembeddedquantization
ToolLiteRT
Coding
google-ai-edge
On-device runtime (formerly TensorFlow Lite) for running models at the edge.
on-devicegpunpuquantization
ToolLlamafile
Deployment
Mozilla-Ocho
Packages an LLM and its runtime into one cross-platform executable file that runs locally with no installation
gguflightweightwhispercross-platform
ToolMonolith
CodingConnectors
tumourlove
MCP plugin for Unreal Engine 5.7 & 5.8 — gives AI assistants full read/write access to Blueprints, Materials, Niagara, Animation, Mesh, AI,
unreal-enginegame-developmentmcp-serverblueprint-scripting
ToolNcnn
Coding
Tencent
ncnn is a high-performance neural network inference framework optimized for mobile, embedded, and desktop deployment.
onnxpytorchtensorflowvulkan
ToolSurogate
TrainingDeployment
invergent-ai
Fine-tuning a small model on a few hundred rows costs a fraction of sending the same task through an API. Soup runs that
fine-tunegrpodpolora
ToolBitNet
Inference
microsoft
Microsoft’s bitnet.cpp is a new CPU-first engine for running 1‑bit LLMs based on the BitNet b1.58 architecture, which us
1-bit-llmquantizationcpu-inferencevector-db
ToolHermes M5Stick Firmware
VoiceInterface
Syax89
M5StickC Plus 2 firmware for a Wi-Fi desk companion that shows Hermes Agent status and sends voice input via Groq STT
m5stickcgroqsttfirmware