Apps / Open WebUI
Managed Open WebUI hosting.
One server, no app limits.
The chat interface for Ollama. Together they are a private ChatGPT on hardware nobody shares.
What it is
Open WebUI is the chat interface for your local models: a polished, multi-user front end for Ollama with conversations, document chat, and model switching, entirely on your server.
Minimum RAM
2 GB
Fits on SC-2GB; SC-4GB is the comfortable pick with room to stack.
What we install
The playbook deploys the current Open WebUI wired to the Ollama beside it, HTTPS forced, signups locked to your people, document upload working for retrieval, and the service supervised.
What managed covers
Platform and server updates, daily snapshots plus dual offsite backups of accounts and chat history, monitoring, and the managed firewall in front of the login.
The app is $0. The managed server is the only bill, and everything above is part of it.
What does managed Open WebUI hosting include?
The interface installed, secured, and connected to your models, with history backed up and access limited to accounts you create. Your team gets the familiar chat experience; the transcript never leaves your server.
Is it actually like ChatGPT to use?
Close enough that nobody needs training: conversations, regenerate, system prompts, model switching mid-chat, and document upload for asking questions against files. The difference is invisible and structural: everything stays on hardware you rent under your name.
Can my whole team use it?
Yes, multi-user with roles is built in: create accounts, restrict models per group, and share the one Ollama underneath. A private company chatbot on a flat server bill, with seat pricing nowhere in sight.
What server does the pair need?
The Ollama underneath sets the requirement: 16 GB for comfortable 7B-class models, with Open WebUI itself adding little. The Private AI kit preselects both and the meter lands you on the right plan honestly.
Stack it
Open WebUI runs well with.
Same server, no extra bill. These are the companions our team installs next to Open WebUI most often.
Ollama
+16 GBRun open LLMs on your own server. Honest requirement: real models want real RAM, and the meter will route you to the right plan.
WireGuard
+1 GBThe modern VPN protocol, deployed with client configs ready to import and the tunnel verified working.
Frequently asked questions
How much RAM does Open WebUI need?
2 GB for the interface itself; the models behind it in Ollama carry the real requirement.
Can it chat with documents?
Yes, upload files and ask questions against them; the retrieval runs locally alongside the model, so the documents stay private too.
Can we keep it off the internet?
Yes: WireGuard from the catalog in front makes the whole AI stack VPN-only, which for client-confidential work is the recommended shape.
Sizing an app stack? The VPS sizing calculator recommends a plan from your apps and traffic, with the math shown.
Shared CPU servers
The right home for Open WebUI and most stacks: 3.0+ GHz vCPU, from $10/mo.
See Shared CPU →Dedicated CPU servers
For CPU-hungry stacks and busy databases: cores that are physically yours.
See Dedicated CPU →All plans and pricing
Every plan, both families, one honest price with everything included.
See pricing →Your server runs. You sleep.
Fully managed hosting from people who have been doing this since 2001.