This is an early release preview. You may encounter bugs.

Catalogue

Submit a tool

Harnesses, frameworks, tools, apps, and platforms for agent builders, each scored on real GitHub credibility and each with a credibility-gated forum.

35 tools · showing 1–24

#openai-compatible ×
Tool

vLLM

Inference

vllm-project

LLM inference and serving library using PagedAttention and continuous batching, with an OpenAI-compatible API server

pagedattentionopenai-compatiblequantization

A+
Tool

AxonHub

CodingInterface

looplj

AI gateway that translates any SDK's requests to any provider, with tracing, RBAC, load balancing, and cost tracking

gatewayobservabilityauthtoken-optimization

A+
Tool

LiteLLM

Inference

BerriAI

Open-source AI gateway exposing 100+ LLM providers through one OpenAI-compatible interface, as a Python SDK or self-hosted proxy

openai-compatibleproxymulti-providergateway

A+
Tool

LocalAI

InferenceGenerative Media

mudler

Self-hosted engine that runs LLM, vision, voice, image, and video models on any hardware behind OpenAI-compatible APIs

local-firstopenai-compatiblellama-cppmultimodal

A+
Tool

Open WebUI

Interface

open-webui

Self-hosted, offline-capable AI platform for Ollama and OpenAI-compatible models with RAG, tools, and multi-user access control

ollamaraglocal-firstopenai-compatible

A+
Tool

FreeLLMAPI

DeploymentCoding

tashfeenahmed

OpenAI-compatible proxy that stacks the free tiers of 16 LLM providers (~1.7B tokens/month) behind one /v1 endpoint — plus any custom OpenAI

gatewayopenai-compatiblefailoversmart-routing

A+
Tool

9Router

Coding

decolua

Unlimited AI coding. Connect Claude Code, Codex, Cursor, Cline, Copilot, Antigravity to free Claude/GPT/Gemini via 40+ providers.

gatewaymulti-providerfallback-routingtoken-optimization

A
Tool

Manifest

Inference

mnfst

Open-source LLM router exposing one OpenAI-compatible endpoint across API keys, subscriptions, and local models, with cost tracking

gatewaytoken-optimizationopenai-compatiblebyok

A
Tool

Jan

InterfaceInference

menloresearch

Desktop app for running open-weight LLMs locally or connecting to cloud providers, exposing an OpenAI-compatible local API

local-firstopenai-compatiblellama-cppdesktop

A
Tool

Smg

Coding

lightseekorg

Engine-agnostic gateway that routes LLM requests across self-hosted and cloud backends with cache-aware load balancing

gatewayvllmopenai-compatible

A
Tool

Mesh LLM

Deployment

Mesh-LLM

Distributed LLM inference that pools GPUs across machines and serves one OpenAI-compatible API, splitting large models across nodes

distributed-inferenceopenai-compatiblegpulocal-first

A
Tool

Mlx Serve

InferenceGenerative MediaVoice

ddalcu

Native LLM inference server for Apple Silicon. OpenAI + Anthropic API compatible. No Python. Includes MLX Core macOS app with chat, agent mo

mlxggufapple-siliconopenai-compatible

A
Tool

Switchyard

Inference

NVIDIA-NeMo

Switchyard lets LLM applications route traffic across models and providers while preserving native OpenAI and Anthropic API compatibility -

gatewaylitellmopenai-compatibletoken-optimization

A
Tool

Apfel

InterfaceInference

Arthur-Ficial

The free AI already on your Mac. CLI tool, OpenAI-compatible server, and interactive chat — all on-device via Apple Intelligence. No API key

apple-intelligenceon-devicemacosopenai-compatible

A
Platform

Exo

InferenceDeployment

exo-explore

Open-source distributed inference runtime that pools everyday devices into one local cluster to serve frontier models behind an OpenAI-compatible API.

distributed-inferencedevice-clusteropenai-compatiblelocal-first

A
Tool

mistral.rs

Coding

EricLBuehler

Rust LLM inference engine with multimodal support, automatic quantization, and OpenAI- and Anthropic-compatible serving

quantizationopenai-compatiblemultimodallora

A
Tool

WebLLM

Inference

mlc-ai

In-browser LLM inference engine running open models client-side on WebGPU with an OpenAI-compatible API

webgpubrowseropenai-compatiblejson-mode

A
Tool

Flama

InferenceInterfaceConnectors

vortico

Ollama is fine for trying a model. vLLM becomes interesting when that model has to serve traffic: many users, agent work

openai-compatiblechatbotasgi

A
Tool

FuXi

CodingInterface

fuxicodex

FuXi is a fast, self-contained AI coding agent that lives in your terminal — edit code, run commands, and drive tools, with cost-aware routi

productivityclicoding-agentopenai-compatible

A
Tool

GoModel

InferenceMonitoring

ENTERPILOT

AI gateway written in Go. Lightweight unified OpenAI-compatible API for OpenAI, Anthropic, Gemini, Groq, xAI & Ollama. LiteLLM alternative w

gatewayproxyopenai-compatibletoken-optimization

A
Tool

AIProxy

CodingMonitoring

labring

AI gateway with OpenAI-, Anthropic-, and Gemini-compatible routing, multi-tenant management, monitoring, and a plugin system

gatewaymulti-tenantopenai-compatiblemonitoring-dashboard

B
Tool

Muna

InferenceDeployment

muna-ai

Compiles Python AI functions into self-contained native binaries and serves open models via an OpenAI-compatible client across cloud, edge and device.

openai-compatibleon-devicegpumodel-compilation

B
Tool

Gateway

Coding

Portkey-AI

Open-source AI gateway routing to 1,600+ LLMs through one API with retries, fallbacks, and guardrails

gatewayopenai-compatibleguardrailsfallback

B
Tool

MLC LLM

Inference

mlc-ai

Machine-learning compiler and engine for deploying LLMs across AMD, NVIDIA, Apple, and Intel GPUs, browsers, iOS, and Android

cross-platformwebgpuon-devicequantization

B