Run Hermes Agent + Claude Code locally on llama.cpp — zero API costs. A 4h / 7M-token session that would have cost $94 on Claude Opus 4.7
-
Updated
Jun 8, 2026 - Shell
8000
Run Hermes Agent + Claude Code locally on llama.cpp — zero API costs. A 4h / 7M-token session that would have cost $94 on Claude Opus 4.7
A self-hosted AI platform — inference, tool use, browser automation, image generation, speech synthesis, transcription, object storage, agentic code execution, and more — behind a single OpenAI-compatible endpoint. One docker-compose up.
Run IBM Granite 4.0 locally on Raspberry Pi 5 with Ollama.This is a privacy-first AI. Your data never leaves your device because it runs 100% locally. There are no cloud uploads and no third-party tracking.
Linux distro with a built-in LLM. Detects your hardware, picks vLLM or Ollama, downloads a model that fits, and serves an OpenAI-compatible API from first boot.
Defense-in-depth platform for running OpenClaw agents on personal hardware
Production-ready guide to serve multiple local LLMs (chat, embeddings, reranking) on NVIDIA DGX Spark using vLLM, an nginx TLS gateway, and systemd automation
Run Codex CLI with self-hosted models in air-gapped networks via LiteLLM.
🐳 One-command self-hosted AI stack — Ollama + Open WebUI + Qdrant + n8n. Your private ChatGPT, fully offline.
Automatisches iCloud Foto/Video Backup auf Synology NAS — Multi-Account, Docker Compose, DSM Aufgabenplaner
Saga robots
One-command self-hosted LobeHub deploy — Docker Compose (PostgreSQL/PGVector, Redis, RustFS S3, SearXNG), optional Cloudflare Tunnel for public access.
Scripts para ejecutar IA local con lógica de memoria personalizada.
Inspired by the idea that programming can be a form of sculpture, each script is a tool for carving time, sound, and color.
One-command, self-hosted OpenClaw starter — Raspberry Pi + Ollama friendly, hardened defaults, curated skills.
Deployment examples for open-source AI models - GPU VMs, Kubernetes, and OpenAI-compatible API
Bash automation for self-hosted Ollama: model management, Open WebUI Docker deployment, and API querying from the command line.
self-organized looped AI examples
Docker-first, local-first AI workload toolkit for macOS Apple Silicon using Ollama, llama.cpp, LiteLLM, and Claude-compatible local endpoints.
Add a description, image, and links to the self-hosted-ai topic page so that developers can more easily learn about it.
To associate your repository with the self-hosted-ai topic, visit your repo's landing page and select "manage topics."