Ramone — Local AI System
A fully self-hosted private AI infrastructure — zero cloud dependency, zero data egress. Five LLMs served via Ollama on local hardware, wrapped in Docker, accessed through Open WebUI with ten specialised workbots backed by RAG knowledge bases. Everything lives on a dedicated NVMe drive: portable, rebuildable in under 30 minutes.
Five local models and ten specialised workbots run without cloud inference or data egress, and the stack remains portable enough to rebuild in under 30 minutes.