Collection
Run agents on your own machine
Runtimes, models and tooling that work on hardware you control.
58 tools · showing 25–48
ToolODS
DeploymentInterfaceInference
Osmantic
Turn your PC, Mac, or Linux box into an AI server. LLM inference, chat UI, voice, agents, workflows, RAG, and image generation.
local-firstollamacomfyuin8n
Atomic Agent
InterfaceMemoryInference
AtomicBot-ai
Local First Ai Agent. Optimized for Local Ai models. Long context window. Proper tools callings. Runs privately on your device.
local-firstllama-cppcomputer-usebrowser-automation
ToolMunder Difflin
CodingMemory
chaitanyagiri
local multi-agent harness
multi-agentlocal-firstdesktoporchestration
ToolOpenMed
Data WranglingSecurity
maziyarpanahi
On-device clinical NLP for biomedical entity extraction and HIPAA PII de-identification, running fully offline from phones to GPU servers
healthcareclinical-nlphipaapiion-device
ToolMagnitude
Inference
magnitudedev
Your fully local, private agent. Runs models on your machine with its built-in inference engine. Works out of the box, on any hardware.
local-firstmodel-managementcli
ToolOpenSquilla
InferenceMemory
TokenRhythm
Token-efficient microkernel AI agent with an on-device model router, persistent memory and pluggable multi-provider support
model-routeron-devicemulti-providercli
ToolMeetily
Voice
Zackriya-Solutions
Self-hosted AI meeting assistant that transcribes, diarizes, and summarizes meetings locally with Whisper/Parakeet and Ollama
productivitywhispersttdiarizationollama
ToolNadirClaw
Inference
NadirRouter
Open-source LLM router & AI cost optimizer. Routes simple prompts to cheap/local models, complex ones to premium — automatically. Drop-in Op
gatewaytoken-optimizationproxylocal-first
ToolOpenCompany
CodingConnectors
zeenie-ai
Self-improving AI that runs your whole business turning LLM tokens into work and dollars.
productivityagent-builderorchestrationwhatsappollama
ToolOGAM
Generative MediaVoiceInterface
off-grid-ai
The Swiss Army Knife of Offline AI. Chat, see, speak, and generate images on your phone or Mac — GGUF LLMs, vision, Whisper speech-to-text,
on-deviceggufwhisperstable-diffusion
ToolMuna
InferenceDeployment
muna-ai
Compiles Python AI functions into self-contained native binaries and serves open models via an OpenAI-compatible client across cloud, edge and device.
openai-compatibleon-devicegpumodel-compilation
ToolDeepAudit
Security
lintsinghua
DeepAudit:人人拥有的 AI 黑客战队,让漏洞挖掘触手可及。国内首个开源的代码漏洞挖掘多智能体系统。小白一键部署运行,自主协作审计 + 自动化沙箱 PoC 验证。支持 Ollama 私有部署 ,一键生成报告。支持中转站。让安全不再昂贵,让审计不再复杂。
code-auditvulnerability-scannergeminiollama
ToolMlx Dspark
Inference
ARahim3
Up to 3× faster LLM decoding on Apple Silicon, lossless. Native MLX port of DeepSeek's DSpark & z-lab's DFlash speculative decoding — Gemma-
speculative-decodingmlxapple-siliconinference-optimization
HarnessOdysseus
InterfaceResearchInference
odysseus-dev
Odysseus is a self-hosted workspace with powerful local tools. Keep auth enabled, keep private data out of Git, and do not expose raw model/service ports publicly.
productivitylocal-firstchatbotdockeremail
PlatformTinyagentos
MemoryDeploymentInference
jaylfc
Self-hosted, framework-agnostic AI agent platform that runs on your own hardware with a browser desktop
local-firstframework-agnosticmulti-frameworkknowledge-graph
ToolEmbedAnything
Data WranglingMemoryInference
StarlightSearch
Rust-based inference, ingestion and indexing library for embeddings and retrieval, with Python bindings.
local-firstcloudragsdk
ToolMLC LLM
Inference
mlc-ai
Machine-learning compiler and engine for deploying LLMs across AMD, NVIDIA, Apple, and Intel GPUs, browsers, iOS, and Android
cross-platformwebgpuon-devicequantization
Dive
Interface
OpenAgentPlatform
mcp-clientdesktopollamaelectron
AppFox in the Box
InterfaceMemory
fox-in-the-box-ai
Self-hosted AI assistant bundling the Hermes agent, chat UI, local memory, and optional Tailscale remote access into one desktop app
ollamatailscalemem0local-first
ToolNeuphonic
VoiceGenerative Media
neuphonic
Text-to-speech via Neuphonic's API.
ttsvoice-cloningon-devicegguf
ToolHiring Agent
Data Wrangling
interviewstreet
Pipeline that scores resumes by extracting structured data from PDFs and enriching it with GitHub signals
recruitingproductivityresume-screeningdocument-processinggithub-signalsollama
ToolMetronix Memory
Memory
mtrnix
Metronix Memory is a self-hosted memory backend for agents over MCP: ingest files and Saa...
ragknowledge-graphlocal-firstollama
ToolHybro Hub
Connectors
hybroai
Daemon that links local AI agents to the hybro.ai portal so local and cloud agents run side by side in one interface
a2amulti-agentlocal-firstollama
ToolMoondream
Inference
m87-labs
Small open-source vision-language model with a local server and client SDKs for VQA, captioning, pointing and object detection in agent pipelines.
vision-language-modelimage-captioningobject-detectionmultimodal