Pre-configured images
Runtime, dependencies, reverse proxy and TLS already set up. No half-day of yak-shaving.
Pre-built, firewalled VPS images for open-source agent runtimes. Your data stays on hardware you control, in a region you chose, with no third-party inference provider in the path.
Control, residency and predictable cost, the three reasons that hold up.
Runtime, dependencies, reverse proxy and TLS already set up. No half-day of yak-shaving.
Only the ports the stack needs are open. Admin interfaces bound to localhost behind an SSH tunnel.
Deploy in the jurisdiction your compliance requirements need.
Not a black box. The configuration is documented, on disk, and yours to change.
No automated backup or snapshot ships with the VPS. Back up your image and data yourself.
You pay for the server. No per-token billing, no seat licences, no usage surprises.
SSD SSD VPS 2 is the practical minimum for most agent runtimes. SSD SSD VPS 8 for anything running local inference.
A great way to get started.
Renews at $17.99/mo
A great way to get started.
Renews at $19.99/mo
A great way to get started.
Renews at $29.99/mo
A great way to get started.
Renews at $49.99/mo
A great way to get started.
Renews at $74.99/mo
A great way to get started.
Renews at $137.99/mo
A great way to get started.
Renews at $259.99/mo
A great way to get started.
Renews at $499.99/mo
KVM SSD VPS, unmanaged, Europe location. Full root access.
This is the step most one-click pages leave out: the moment after you have picked an image and the server exists, and you are looking at it wondering what to do next. Here is the complete path, with nothing assumed.
That is a fair thing to find offputting, and it is why we say plainly in the FAQs below that this product is aimed at developers and technical teams rather than a no-code audience. If you want the same category of tool without touching a terminal, ask our pre-sales team, we can point you at a managed alternative rather than sell you a server you will not enjoy configuring.
Each image is maintained against upstream releases and rebuilt when a security update lands. The image version and its upstream commit are recorded in /etc/kryohost-image so you always know what you are running.
| Image | What it is | Minimum plan | Notes |
|---|---|---|---|
| OpenClaw | Open-source agent runtime with a tool-calling loop | SSD VPS 2 | Bring your own model API credentials |
| NousResearch Hermes | Open-weights assistant model server | SSD VPS 8 | Local inference; GPU strongly recommended |
| Ollama | Local model runner with a simple HTTP API | SSD VPS 8 | CPU inference is slow but functional for small models |
| n8n | Workflow automation with LLM nodes | SSD VPS 2 | Good bridge between agents and existing systems |
| LangChain starter | Python environment with LangChain and FastAPI | SSD VPS 2 | A scaffold, not an application |
| Vector DB (Qdrant) | Vector database for retrieval-augmented generation | VPS 4 | Pairs with any of the above |
All images ship with unattended security updates enabled, SSH key authentication only, and a firewall denying inbound traffic by default. What they do not ship with is an opinion about your application architecture, they are starting points.
You can run small open-weights models on a CPU-only VPS, and people do. Expect a few tokens per second on a 7B model at SSD VPS 8, usable for batch work, frustrating for anything interactive. If you need interactive local inference, you need a GPU, and we would rather tell you that before you buy than after.
Every conversation about self-hosted AI agent hosting eventually comes back to hardware, and for good reason. You can tune a stack endlessly, but you cannot compensate in software for a disk that is queueing, a CPU that is oversubscribed four times over, or an upstream carrier that routes your visitors halfway around the planet before delivering the first byte. KryoHost builds from the metal upward precisely because the metal sets the ceiling on everything above it.
Our our agent images fleet runs on dual-socket AMD EPYC and Intel Xeon Scalable nodes with ECC memory, paired exclusively with enterprise NVMe drives in RAID-10. There is no SATA tier hiding behind the marketing copy, and no "NVMe-accelerated" wording that quietly means a small cache in front of spinning disks. NVMe changes the character of a server rather than simply making it faster: random read latency drops from milliseconds to microseconds, so the database queries that dominate page-generation time on self-hosted AI agents, inference endpoints and automation stop being the bottleneck. In our own benchmarks, moving an unchanged WordPress install from a SATA SSD node to an NVMe node cut median time-to-first-byte by 38% without a single line of code being touched.
Capacity planning is the unglamorous half of the story. We cap node density well below what the hardware could theoretically carry and we alarm on sustained CPU steal, disk queue depth and memory pressure rather than waiting for customers to open tickets. When a node crosses its threshold, new provisioning stops on that node and workloads are rebalanced. That is why the phrase "the server was fine, your site is just heavy" is one you will not hear from our team, if steal time is climbing, that is our problem to solve, not yours to absorb.
KryoHost operates across Tier-III and Tier-IV facilities in 47 customer-selectable countries. Each core site is multi-homed across at least three Tier-1 carriers and connected to the dominant regional internet exchange, so traffic reaches your visitors through the shortest sensible path rather than the cheapest available one. Blended transit is convenient for a provider and mediocre for a customer; direct peering costs more and is the reason a visitor in Frankfurt does not have their packets tour Europe before they see your homepage.
The practical effect of all this is measurable rather than decorative. Our published 99.99% uptime SLA is backed by service credits, and our public status page records every incident, including the ones that lasted four minutes and that nobody noticed. A provider that only publishes its good quarters is not publishing anything useful.
Most sites are not compromised by a determined attacker studying them for weeks. They are compromised by automated scanners walking the entire IPv4 space, trying known plugin vulnerabilities and credential lists against everything that answers on port 443. Defence against that reality is layered and mostly boring, which is exactly why it works, and why it should already be switched on when your deployment is delivered rather than sold as an upsell after the incident.
Website files on every deployment are backed up twice a week, with email accounts backed up on a separate schedule so both are protected without either job blocking the other. You can also generate and download a full account backup at any time from cPanel, and our support team can help with a restore if you need one.
A twice-weekly schedule is not the same as continuous protection, and we would rather you knew that than assumed otherwise. If you publish or sell something several times a day, keep your own more frequent backup for the hours between our scheduled runs, particularly for a database that changes constantly, such as an order table.
Compliance-wise, our infrastructure supports GDPR data-residency requirements through EU-only regions, and we sign Data Processing Agreements on request. For workloads touching payment data, our environment is PCI-DSS ready, you still own the compliance of your own application, but the platform underneath will not be the reason an assessment fails.
Search engines have been explicit that page experience is a ranking input, and Core Web Vitals are the measurable part of that signal. What gets lost in most self-hosted AI agent hosting marketing is which of those metrics hosting can genuinely influence. Being precise about it saves you money, because it tells you when to buy a bigger deployment and when to fix your front end instead.
| Metric | What it measures | Hosting influence | What actually fixes it |
|---|---|---|---|
| TTFB | Time to first byte from the server | Very high | Faster CPU/disk, server-side caching, closer region, HTTP/3 |
| LCP | Largest contentful paint | High | Low TTFB, image compression, CDN delivery, preloading the hero asset |
| CLS | Cumulative layout shift | None | Width/height attributes on media, reserved ad slots, font-display strategy |
| INP | Interaction to next paint | Low | Less blocking JavaScript, smaller third-party bundles, deferred scripts |
| FCP | First contentful paint | Medium | TTFB plus render-blocking CSS removal and critical-CSS inlining |
Read that table honestly and a useful rule emerges: hosting owns the first 200–400 milliseconds and your front end owns most of what follows. That is not an argument for cheap hosting (those first milliseconds are a floor that nothing downstream can undo) but it is an argument against believing a plan upgrade will rescue a page carrying eleven tracking scripts and a 4 MB uncompressed hero image.
KryoHost our agent images ships with a caching stack that is configured on day one rather than left as an exercise. Requests are answered at the shallowest layer that can serve them correctly, and each layer that handles a request is one your origin never sees.
Compression and protocol choices are handled at the edge too. Brotli is preferred over gzip where the client supports it, HTTP/3 with QUIC is enabled by default so lossy mobile networks stop paying the TCP head-of-line penalty, and TLS 1.3 removes a full round trip from the handshake. None of these require a support ticket; they are the default configuration on every deployment.
A default WordPress install with a commercial theme, tested from a cold cache in the same region as the server, consistently returns in 180–260 ms TTFB on our NVMe nodes. With the full-page cache warm, that drops to 40–70 ms. We publish the ranges rather than the single best run, because the best run is the number every host quotes and none of them reproduce.
Support is where hosting companies quietly differ most, and where the difference is hardest to evaluate before you buy. Every provider claims 24/7 availability. Far fewer will tell you who is on the other end at 03:00 on a Sunday, how many tickets that person is holding, or what percentage get resolved without being escalated into a queue you cannot see.
KryoHost staffs Linux system administrators across three timezones. There is no offshore first line whose job is to send you a knowledgebase link and close the ticket. Our commitment is a first response inside < 15 min on anything service-affecting, escalating automatically to on-call if that is missed. When something does need a second pair of hands it moves to a named senior engineer and you are told who owns it, not dropped into a silent queue.
Being clear about scope up front prevents the most common support frustration: discovering after an incident that the thing you assumed was covered never was.
| Request | Included | Notes |
|---|---|---|
| Server, network and platform faults | Yes | Always our responsibility, always free |
| Migration from another host | Yes | Unlimited sites on standard control panels |
| Email deliverability (SPF, DKIM, DMARC) | Yes | We configure the records and verify alignment |
| SSL installation and renewal | Yes | Automatic for free certificates, assisted for third-party |
| CMS core, plugin and theme updates | Yes | On managed WordPress plans |
| Malware cleanup after a compromise | Yes | One free deep clean per year, per account |
| Performance triage and cache tuning | Yes | We will tell you honestly if the fix is in your code |
| Writing or debugging your application code | No | We will point at the failing query or function |
| Custom theme and design work | No | Available through our partner network |
| Third-party SaaS integrations | No | We support the server side of the connection |
Support reaches you through live chat, tickets and email, and (on business and VPS plans) a scheduled screen-share call when something genuinely needs a live conversation. The same standard of engineer answers at any hour, which matters more than it sounds when you are describing an intermittent fault under pressure at two in the morning.
Buying hosting badly usually happens in one of two directions. Either you buy the cheapest thing available, hit its ceiling within a quarter and migrate under pressure; or you buy three tiers above what you need and pay for headroom that sits idle for years. Both are avoidable with about ten minutes of honest arithmetic.
Storage is the number hosting companies advertise because it is cheap to give away and easy to compare. It is almost never the constraint. A well-built content site with a few hundred posts occupies under 2 GB. What actually determines whether you outgrow a deployment is concurrency, how many requests need generating simultaneously at your busiest minute, not your busiest month.
Take your peak day, look at the busiest hour, and assume the busiest minute carries roughly three times the average minute in that hour. If that number is under about 60 dynamic requests, entry our agent images is genuinely sufficient. Between 60 and 300, you want guaranteed CPU. Above that, you want a VPS or a container-isolated business plan where noisy neighbours cannot reach you at all.
| Your situation | Sensible starting point | What triggers the next upgrade |
|---|---|---|
| Personal blog or portfolio, under 10k visits/month | Entry shared plan | Adding a store or a membership area |
| Small business site with a contact form | Entry or mid shared plan | Consistently exceeding 25k visits/month |
| Content site publishing weekly, 25k–150k visits | Mid shared or managed WordPress | Admin feeling slow while traffic is fine |
| WooCommerce store under 500 SKUs | Business hosting | Checkout slowdowns during promotions |
| Store over 500 SKUs or heavy filtering | VPS 2 or Business Max | Sustained CPU above 70% |
| SaaS, API or custom application | VPS 2 to VPS 8 | Needing horizontal scale or HA |
| Agency hosting client sites | Reseller | Passing roughly 120 active accounts |
The introductory rate on a self-hosted AI agent hosting plan is a customer-acquisition cost, not a price. What you will actually pay is the renewal figure, multiplied by however many years you keep the site. We publish both on every plan card on this page for exactly that reason. A $19.99 entry price that renews at four times the rate is not cheap hosting; it is a loan against your future attention.
Compare on total cost over three years, including domain renewal, SSL if it is not free, backups if they are billed separately, and any migration fee at the point you eventually leave. That last one matters: a host that charges to let you go has told you something about its confidence in the product.
Hosting vocabulary is unusually good at making simple ideas sound complicated, and occasionally at making limited products sound generous. Here is what the terms on a self-hosted AI agent hosting comparison page actually mean in practice.
| Term | What it means in practice |
|---|---|
| Unmetered bandwidth | No hard transfer cap, but a fair-use policy applies. Normal sites never approach it; a file-distribution service will. |
| Unlimited storage | Unlimited for ordinary website files. Backup archives, media libraries and mail spools usually have separate soft caps in the terms. |
| NVMe | Flash storage on the PCIe bus rather than a SATA cable. Several times the throughput and a fraction of the latency of a SATA SSD. |
| vCPU | A virtual core mapped to a physical thread. A dedicated vCPU is reserved for you; a shared vCPU competes with neighbours. |
| TTFB | Time to first byte, how long the server takes to start answering. The clearest single indicator of hosting quality. |
| LiteSpeed | A web server that is API-compatible with Apache but handles concurrency with an event-driven model, plus a built-in full-page cache. |
| Object cache | A memory store (Redis or Memcached) holding query results so the database is not asked the same question repeatedly. |
| CDN | A network of edge servers caching your static assets close to visitors, cutting both latency and origin load. |
| SLA | A contractual uptime commitment with defined compensation. Without service credits attached, it is a marketing number. |
| Soft limit | A threshold that throttles rather than stops you. Read these carefully, they are where "unlimited" actually ends. |
| CageFS | Per-account filesystem virtualisation. It is what stops one compromised account on a shared node reading another's files. |
| Steal time | CPU cycles your VM wanted but the host gave to someone else. Persistent steal time means the node is oversold. |
If a provider will not define its soft limits in writing, treat that as the answer. Ours are published in the Terms of Use, in plain language, with the numbers included.
Billing should be the least interesting part of your relationship with a hosting provider. It becomes interesting only when a provider uses it as a retention mechanism, automatic multi-year renewals, cancellation flows hidden behind three menus, refunds that require a phone call during business hours in a timezone you do not live in. We have deliberately built ours to be dull.
We accept cards, PayPal, bank transfer and major cryptocurrencies. Invoices are issued with full tax details for business accounts, and annual billing carries a genuine discount rather than a discount against an inflated monthly rate.
On the SLA: 99.99% availability is measured monthly at the network edge, excluding scheduled maintenance announced at least 72 hours in advance. Fall below it and you receive service credits automatically, you do not have to notice, calculate and claim. An SLA that requires the customer to police it is not much of a commitment.
Still unsure? Our team answers pre-sales questions 24/7, usually in under 15 minutes.
You select the image at checkout and the server arrives with the runtime installed, dependencies resolved, a reverse proxy configured with TLS, a firewall permitting only what the stack needs, and admin interfaces bound to localhost. What it does not mean is that we have made configuration decisions you cannot see or change, everything is on disk, documented, and yours. It removes the setup afternoon, not your control.
No, and this is the most important thing to understand before you buy. We provide the infrastructure and the runtime. You supply either API credentials for a hosted model provider, or open-weights models you download onto the server. If you intend to run inference locally rather than calling an API, size accordingly, that is a fundamentally different hardware requirement and CPU-only inference on a small VPS will disappoint you.
For an agent runtime calling a hosted model API, SSD SSD VPS 2 with 4 GB is comfortable, the workload is mostly orchestration and I/O rather than computation. For local inference on small open-weights models, SSD VPS 6 with 32 GB is the practical floor and performance will still be modest without a GPU. If local inference is central to what you are building, talk to us about GPU-attached instances before ordering a standard VPS.
On the server, yes, it is your machine, in your chosen region, and we do not inspect customer workloads. But be precise about the whole path: if your agent calls a third-party model API, your prompts go to that provider under their terms, and self-hosting the orchestration layer does not change that. Genuine end-to-end privacy requires local inference. We would rather state that plainly than let the word "self-hosted" do misleading work.
Provisioning is automatic and normally completes in under 60 seconds. You receive your control panel login, nameservers and connection details by email the moment the deployment is live. Orders paid by bank transfer activate once the payment clears, which usually takes one to two business days.
Yes. Upgrades within the same product family are applied in place, take effect within minutes and are billed pro-rata against your remaining term. Downgrades take effect at your next renewal date, provided your current usage fits within the smaller plan's limits. Moving between product families (shared to VPS, for example) is handled by our migration team at no charge.
It is, with no limit on the number of sites for standard control-panel migrations. Our team handles the copy, verification and DNS cutover, keeps your existing site serving traffic throughout, and retains a rollback copy for 14 days after the switch. Complex custom-stack migrations are quoted individually, and we will tell you before any work starts if yours falls into that category.
You are notified by email well before anything is enforced, with the specific metric and figure included. Brief spikes (a post going viral, a promotion landing) are absorbed rather than punished. Only sustained overuse leads to throttling, and we will always propose a right-sized plan before that point. Nothing is suspended without prior written notice except in cases of active abuse.
Every domain and subdomain on every plan gets a free DV certificate, issued automatically at setup and renewed automatically for as long as the site is hosted with us. Paid OV, EV and wildcard certificates are available if you need a warranty, organisation validation or a single certificate covering unlimited subdomains.
We operate in 47 customer-selectable countries across Europe, Asia, North America, South America, Africa and Oceania. You pick your region at checkout, and you can relocate later at no cost. Choose the region closest to the majority of your visitors, it is the single cheapest performance improvement available to you.
Our published SLA is 99.99% measured monthly at the network edge, and it is backed by automatic service credits rather than a claims process. Historical uptime, including every incident and its duration, is on our public status page, the bad months are there alongside the good ones.
Standard VPS pricing and hourly billing. Trying one costs less than lunch.
Free migration · No setup fees · Cancel any time