This is an early release preview. You may encounter bugs.

Run agents on your own machine

Runtimes, models and tooling that work on hardware you control.

58 tools · showing 25–48

Tool

ODS

DeploymentInterfaceInference

Osmantic

Turn your PC, Mac, or Linux box into an AI server. LLM inference, chat UI, voice, agents, workflows, RAG, and image generation.

local-firstollamacomfyuin8n

A
Tool

Atomic Agent

InterfaceMemoryInference

AtomicBot-ai

Local First Ai Agent. Optimized for Local Ai models. Long context window. Proper tools callings. Runs privately on your device.

local-firstllama-cppcomputer-usebrowser-automation

A
Tool

Munder Difflin

CodingMemory

chaitanyagiri

local multi-agent harness

multi-agentlocal-firstdesktoporchestration

A
Tool

OpenMed

Data WranglingSecurity

maziyarpanahi

On-device clinical NLP for biomedical entity extraction and HIPAA PII de-identification, running fully offline from phones to GPU servers

healthcareclinical-nlphipaapiion-device

A
Tool

Magnitude

Inference

magnitudedev

Your fully local, private agent. Runs models on your machine with its built-in inference engine. Works out of the box, on any hardware.

local-firstmodel-managementcli

A
Tool

OpenSquilla

InferenceMemory

TokenRhythm

Token-efficient microkernel AI agent with an on-device model router, persistent memory and pluggable multi-provider support

model-routeron-devicemulti-providercli

A
Tool

Meetily

Voice

Zackriya-Solutions

Self-hosted AI meeting assistant that transcribes, diarizes, and summarizes meetings locally with Whisper/Parakeet and Ollama

productivitywhispersttdiarizationollama

A
Tool

NadirClaw

Inference

NadirRouter

Open-source LLM router & AI cost optimizer. Routes simple prompts to cheap/local models, complex ones to premium — automatically. Drop-in Op

gatewaytoken-optimizationproxylocal-first

A
Tool

OpenCompany

CodingConnectors

zeenie-ai

Self-improving AI that runs your whole business turning LLM tokens into work and dollars.

productivityagent-builderorchestrationwhatsappollama

A
Tool

OGAM

Generative MediaVoiceInterface

off-grid-ai

The Swiss Army Knife of Offline AI. Chat, see, speak, and generate images on your phone or Mac — GGUF LLMs, vision, Whisper speech-to-text,

on-deviceggufwhisperstable-diffusion

B
Tool

Muna

InferenceDeployment

muna-ai

Compiles Python AI functions into self-contained native binaries and serves open models via an OpenAI-compatible client across cloud, edge and device.

openai-compatibleon-devicegpumodel-compilation

B
Tool

DeepAudit

Security

lintsinghua

DeepAudit:人人拥有的 AI 黑客战队,让漏洞挖掘触手可及。国内首个开源的代码漏洞挖掘多智能体系统。小白一键部署运行,自主协作审计 + 自动化沙箱 PoC 验证。支持 Ollama 私有部署 ,一键生成报告。支持中转站。​让安全不再昂贵,让审计不再复杂。

code-auditvulnerability-scannergeminiollama

B
Tool

Mlx Dspark

Inference

ARahim3

Up to 3× faster LLM decoding on Apple Silicon, lossless. Native MLX port of DeepSeek's DSpark & z-lab's DFlash speculative decoding — Gemma-

speculative-decodingmlxapple-siliconinference-optimization

B
Harness

Odysseus

InterfaceResearchInference

odysseus-dev

Odysseus is a self-hosted workspace with powerful local tools. Keep auth enabled, keep private data out of Git, and do not expose raw model/service ports publicly.

productivitylocal-firstchatbotdockeremail

B
Platform

Tinyagentos

MemoryDeploymentInference

jaylfc

Self-hosted, framework-agnostic AI agent platform that runs on your own hardware with a browser desktop

local-firstframework-agnosticmulti-frameworkknowledge-graph

B
Tool

EmbedAnything

Data WranglingMemoryInference

StarlightSearch

Rust-based inference, ingestion and indexing library for embeddings and retrieval, with Python bindings.

local-firstcloudragsdk

B
Tool

MLC LLM

Inference

mlc-ai

Machine-learning compiler and engine for deploying LLMs across AMD, NVIDIA, Apple, and Intel GPUs, browsers, iOS, and Android

cross-platformwebgpuon-devicequantization

B
Tool

Dive

Interface

OpenAgentPlatform

mcp-clientdesktopollamaelectron

B
App

Fox in the Box

InterfaceMemory

fox-in-the-box-ai

Self-hosted AI assistant bundling the Hermes agent, chat UI, local memory, and optional Tailscale remote access into one desktop app

ollamatailscalemem0local-first

B
Tool

Neuphonic

VoiceGenerative Media

neuphonic

Text-to-speech via Neuphonic's API.

ttsvoice-cloningon-devicegguf

B
Tool

Hiring Agent

Data Wrangling

interviewstreet

Pipeline that scores resumes by extracting structured data from PDFs and enriching it with GitHub signals

recruitingproductivityresume-screeningdocument-processinggithub-signalsollama

B
Tool

Metronix Memory

Memory

mtrnix

Metronix Memory is a self-hosted memory backend for agents over MCP: ingest files and Saa...

ragknowledge-graphlocal-firstollama

B
Tool

Hybro Hub

Connectors

hybroai

Daemon that links local AI agents to the hybro.ai portal so local and cloud agents run side by side in one interface

a2amulti-agentlocal-firstollama

C
Tool

Moondream

Inference

m87-labs

Small open-source vision-language model with a local server and client SDKs for VQA, captioning, pointing and object detection in agent pipelines.

vision-language-modelimage-captioningobject-detectionmultimodal

C

More ways in