Ollama Hosting Live In 5 Minutes

Deploy Ollama on infrastructure you control, instead of sending every prompt to a hosted API. xCloud installs the runtime as a One-Click App on a Docker + NGINX server, validates disk space before anything downloads, and publishes it behind HTTPS on your domain — ready for you to pull a model from the dashboard.

  • No terminal
  • No technical skills needed
  • No lock-in, cancel anytime

From Sign-Up to a Running Model In Just 3 Easy Steps

Self-hosting an LLM runtime normally means a container, a reverse proxy, a domain and a certificate before a single model is pulled. This is the same result without assembling any of it by hand.

  1. Step 01

    Sign Up

    Create your xCloud account, spin up a Docker + NGINX server, and choose Ollama as the App Type. That choice is what tells xCloud this is an LLM server — pick General or Premium, a region and an Ubuntu version.

  2. Step 02

    We Deploy It For You

    One click. We bring up Ollama on the Docker + NGINX stack, validate free disk space before anything installs, and publish it over HTTPS on your domain — no SSH session required.

  3. Step 03

    Pull A Model And Call The API

    Open Ollama → Management, click Add New Model, choose one from the dropdown and pull it. Once it's resident, call the HTTP API from your own application or point a chat UI at it.

Everything Ollama Does, On Your Own Server

Ollama runs in full here — model downloads, the HTTP API and lifecycle management, on infrastructure you control.

  • Ollama's Full Model Library

    Pull any model in Ollama's own library that fits your server's memory, straight from the dropdown in Ollama → Management.

  • HTTP And OpenAI-Compatible API

    Call the runtime from your own software over its documented HTTP API, including an OpenAI-compatible route for repointing an existing integration.

  • One-Click Model Management

    Add or remove models from the dashboard after deployment — no terminal session, no manual download.

  • A Chat UI When You Want One

    Pair it with Open WebUI or LibreChat, both one-click apps, when people rather than programs need to talk to the model.

  • Point An Agent At Your Own Model

    Expose the API through your server's NGINX config and add it as a model provider in OpenClaw, for agent workflows with no per-token API cost.

  • Disk Space Validated Before Install

    Free space is checked before a containerized app installs, so a multi-gigabyte model download can't fail halfway for running out of room.

Everything You Need To Run Your Private Inference 24/7

Managed hosting for Ollama, so the parts that keep it online are not your problem.

  • Ollama deployed instantly
  • HTTP and OpenAI-compatible API included
  • Pair it with Open WebUI or LibreChat for your team
  • Free SSL certificate
  • Server backups and snapshots
  • Server security updates included
  • 30+ server locations
  • Unused balance refunded

Simple Pricing. Everything Included.

Dedicated Cloud VPS sizes, pre-configured for Ollama. Pick the one that fits the model you plan to run, and resize any time.

Cloud VPS 6 GB

Below the RAM most models need once xCloud's 20-30% headroom rule is applied — fine for a quick test.

$24.99per month

  • 6 GB RAM
  • 4 vCPU cores
  • 100 GB NVMe SSD
  • 30 TB
  • Unlimited
Get started

Cloud VPS 16 GB

Most popular

Real headroom for a small-to-mid-sized model, its context, and the operating system.

$59.99per month

  • 16 GB RAM
  • 6 vCPU cores
  • 200 GB NVMe SSD
  • 30 TB
  • Unlimited
Get started

Cloud VPS 24 GB

For a larger model, or Ollama running alongside another workload on the same server.

$84.99per month

  • 24 GB RAM
  • 8 vCPU cores
  • 300 GB NVMe SSD
  • 30 TB
  • Unlimited
Get started

Every size is a dedicated Cloud VPS we provision, patch, monitor and restore. Billed monthly with no lock-in — delete the server and the unused balance comes back to your account. Ollama itself is free under the MIT License. The largest published models need far more than 24 GB; connect your own GPU machine through xCloud for those.

Already Have a Server? Bring Your Own

Connect any VPS or GPU machine — DigitalOcean, Vultr, Linode, AWS, Hetzner or bare metal — and deploy Ollama on it through the same panel, for a per-server management fee instead of a Cloud VPS rental.

Free

Start your journey without cost.

$0forever

  • 1 server
  • 10 sites
Start free

Starter

Most popular

Your rate with 2–5 servers connected.

$5per server / month

  • 2–5 servers
  • Unlimited sites
  • All features included
  • Bring any VPS provider
Get started

Professional

The same platform, discounted at 6–10 servers.

$4per server / month

  • 6–10 servers
  • Unlimited sites
  • All features included
  • Bring any VPS provider
Get started

Agency

The same platform, discounted at 11+ servers.

$3per server / month

  • 11+ servers
  • Unlimited sites
  • All features included
  • Bring any VPS provider
Get started

Priced per connected server, so your rate adjusts automatically as you add or remove servers. You pay your own provider for the server itself — the only path to a GPU machine, since the managed ladder above is CPU.

Ready To Run Your Own Private AI Model?

Ollama on xCloud runs on a dedicated Cloud VPS that we provision, patch, monitor and restore. Deploy faster. Own the server. Keep every prompt off someone else's API.

Hear what people are saying about us

See why agencies, developers, startups, and growing businesses choose xCloud to power their projects.

xCloud works really well, I am surprised by its performance being this the release version, I have tried many paid platforms such as runcloud, cloudpages or serveravatar, as well as free platforms such as CyberPanel, CloudPanel or Hestia Control Panel and direct competitors such as SpinupWP and FlyWP. Of all those that I have tried, the platform that works best for WordPress is SpinupWP and I believe that xCloud can reach and surpass that development, its roadmap is promising. Congratulations to the xCloud team, I hope they become the leading platform in the segment, it has all the potential to achieve it!
Brian Geovanny
I've been with runcloud for the last 3 years. However, after trying xcloud for the last month, I'm switching over. They are on a roll with their setup.
Kevin Pineda
I'm moving my servers to xCloud, a cloud control panel by M Asif Rahman and the WPDeveloper team. Their roadmap looks interesting and no, there are no team mates nor API limits with xCloud! #xCloud #wordpress #hosting
Leo Koo
How Cloud Nine Web Elevated Client Experience While Streamlining Hosting Operations using xCloud
Ed Ellingham
xCloud is like having an expert server team on call — without the big agency costs. It's given me back my time and given my clients peace of mind. Highly recommended for freelancers and small teams who want reliable performance and security.
Dumitru Butucel
xCloud isn't just a solution. It's like one of that friend who got your back when you screw things up pretty bad.
Waqar Ahmed
Trust me when I say, once people explore the potential of xCloud they gonna switch from other panels, just keep delivering the best tech, lots of love for the team 😍😍
Rahul Singh
So far my experience has been smooth like silk. I didn't imagine that configuring and managing a VPS would be so easy. As of today, I have deployed 8 sites. The support is good, and they treat users with respect, even if they ask stupid questions (myself included). For the future, I hope for greater flexibility in scheduled backups and more comprehensive and consistent documentation. I'm sure that within a few months, xCloud will become the leader in its segment. This community is amazing, and the Staff here is super helpful.
Massimo Villa
Came from cloudways then to runcloud then to xcloud. Looks like I'm staying here.
Gerson El
I've recently taken all of my Vultr-hosted sites and gotten rid of cPanel for them to substitute it with xCloud. I have to say, I'm really impressed with xCloud so far! The interface is definitely the standout feature for me.
Tony Lewis
I came from Cloudways and won't go back. Highly impressed with xCloud. Glad I gave them a chance.
Jared McDowell
I like xCloud better than RunCloud, but one thing I'll point out is that some years ago RunCloud was hungry for business, the support was outstanding, and the community was active and helped each other. Maybe it was sold, there was a change in management, or they stopped being hungry, but the quality of the RunCloud offering declined and it hasn't been moving forward quickly. My point is that currently the xCloud team has the edge and it is theirs to keep or lose
David McCan
How Kevin Heinrichs Cut Hosting Costs by 84% While Running WooCommerce & Node.js Sites Efficiently
Kevin Heinrichs
I had 150 sites on cloudways and am in the process of moving them to xcloud. Much faster and better priced.
Dan Spinoza
Dear xCloud Team, I have been using xCloud for some time now and am genuinely impressed with its capabilities and the flexibility it offers by allowing users to bring their own VPS for management. It's a fantastic service, and my experience so far has been very positive. Successful migrated 5 sites on a single server. Xcloud migration tool worked great. There were few hiccups with security plugin like MalCare but disabling the plugin prior to migration worked just fine.
Rahil Maknojia
Came from runcloud. Tried xcloud and never came back.
Karl Fanega
How Advanced Web Strategies Saved Over $2,000 & Recovered 56 Hours per Month Managing 70+ Client Websites
Gary Moody
xCloud gave me what no other hosting platform could: the flexibility of a VPS with the convenience of managed WordPress. As a solo founder running an entire business on one server, that combination is an important unfair advantage.
Jorge de los Reyes Martínez
I use Server Avatar and xCloud and I'll sooner use xCloud than Server Avatar personally.
Jay Pagano
How Newman Web Solutions Ensured Perfect Uptime and Seamless Scaling For 120+ WordPress Sites with xCloud
Lori Newman
I too have moved over to xCloud. It's pretty sweet. One click n8n and uptime Kuma apps
John Aperture
I went from Cloudways to xCloud. No intention on switching in the foreseeable future.
Blaine Moore

Frequently Asked Questions

Feel free to contact us with any questions or concerns you may have.

What is Ollama?

Ollama is an open-source runtime for open large language models — it downloads a model, keeps it resident in memory, and answers requests over an HTTP API. This page deploys it as a One-Click App on a managed server.

Is Ollama free?

Yes — Ollama itself is free and open source under the MIT License. What you pay xCloud for is the dedicated server it runs on and the management around it, not the software itself.

Can I call it from my own application?

Yes. The runtime exposes an HTTP API, including an OpenAI-compatible route, so an existing integration can usually be repointed at your server by changing a base URL and a key. Put a credential in front of it first — Ollama does not authenticate callers itself.

Which models can I run?

Anything in Ollama's own library that fits in your server's memory, pulled from the dropdown in Ollama → Management. The catalogue is Ollama's, so it grows when they publish new models rather than when we do — we do not certify a model as supported until its installation and runtime behavior are verified.

How do I add or remove a model after deployment?

Ollama → Management in the dashboard. Add New Model opens a dropdown and pulls the one you choose; the bin icon next to a model removes it.

Why does it need a Docker and NGINX server?

Ollama is a containerized app, so that stack is required rather than preferred. xCloud preselects it as soon as you choose Ollama as the App Type, which is also why it can't be added to an existing WordPress-style server.

Do models download again every time the server restarts?

No. Weights are written to the server's disk on first pull and stay there, which is why xCloud validates free disk space before a containerized app installs.

How fast will it be?

We have not published tokens-per-second numbers for specific server tiers, because we have not finished benchmarking them and an estimate would be worse than saying so. What's true today: CPU inference is slower than GPU, and larger models are slower than smaller ones on the same hardware.

CPU or GPU?

CPU works for batch work and a small team chatting at reading pace. Anything interactive, concurrent, or built on the largest models wants a GPU — xCloud's own self-hosted LLM guide recommends a GPU server for optimal performance. The managed ladder here is CPU; for a GPU, connect a machine from your own provider and manage it through xCloud the same way.

Is my data private?

Prompts, documents and completions stay on the server you control, since nothing needs to leave it to produce an answer. Access to that server — firewall rules, credentials, who can reach the API — is yours to configure.

Can I run a chat UI for my team?

Yes, and there are two. Open WebUI is a straightforward private chat window; LibreChat is a fuller platform, useful when the same window needs to reach hosted providers as well as your local model. Both are one-click apps, and both installers can bring Ollama up with them.

Can an AI agent use my self-hosted model instead of a paid API?

Yes. Expose the Ollama API through the server's NGINX config, then add that endpoint as a model provider in OpenClaw through its own UI rather than by hand-editing the config file.

What should I back up?

Pulled model weights and any application data from a UI you've paired with it (Open WebUI or LibreChat), together and from the same point in time.

Does xCloud automatically back up or update Ollama?

No. Backups run on the schedule you configure against your own S3-compatible store, and moving Ollama or a model to a newer version is a deliberate action you take, not something that happens automatically.

What are the server requirements, and why does this page sell 16 GB?

Ollama itself has almost no footprint; the constraint is the model you choose. xCloud's own guidance is to keep at least 20 to 30 percent more RAM than the model's weight size, and even a small model's weights (roughly 4-5 GB for common 4-bit builds) already push past a 6 GB server once that headroom and the OS are accounted for. This page sells the 16 GB tier as the entry rung so that floor isn't the ceiling on day one.

Is xCloud officially partnered with Ollama?

No. No official partnership or endorsement between xCloud and the Ollama project has been verified. This is a deployment and server-management option for the open-source Ollama runtime.

Can I cancel anytime?

Yes. There is no lock-in — cancel any time. Hosting is billed monthly and charged for what you actually use, so when you delete the server the unused balance is refunded to your account, less up to 10% in payment-processing fees.

Ready to Run Your Own Private AI Model?

Connect your first server in minutes. No credit card required, and no plan you have to grow into.