Managed AnythingLLM VPS hosting
AnythingLLM, live in minutes.
On a server that is yours alone.
Chat with your own documents. A private RAG workspace that ingests PDFs and sites, pairs with Ollama or any API model, and keeps data on your server.
- 500% SLA
- 30-day money-back
- Free Daily Snapshots
- Firewall & IDS
What it is
AnythingLLM turns your documents into something you can talk to: a private RAG workspace that ingests PDFs, docs, and websites, embeds them, and answers questions grounded in your own material, with citations.
Minimum RAM
4 GB
Fits on SC-4GB; SC-8GB is the comfortable pick with room to stack.
What we install
Our playbook deploys the official Docker image with persistent storage for your documents and vector data, HTTPS issued and forced, and multi-user mode ready. Point it at the model you prefer: your API key or an Ollama on the same server.
Every plan includes
Instant hostname with TLS, database created and wired, daily snapshots, firewall and IDS. Add Fully Managed anytime and our engineers run the whole server for you.
Pick your server
Choose a plan for AnythingLLM.
AnythingLLM needs at least 4 GB of RAM. Every plan is your own private VPS, and the same server runs as many apps as fit.
SC-4GB
- 2 vCPU
- 4 GB RAM
- 100 GB NVMe
- 10 TB transfer
- Full Root Access
- Non-Oversold Cores
- 100% NVMe Storage
- Network Firewall/IDS
- Free Daily Snapshots
- DDoS-Protected Network
- Dual Backups (Optional)
- Fully Managed (Optional)
30-day money-back
SC-6GB
- 3 vCPU
- 6 GB RAM
- 150 GB NVMe
- 15 TB transfer
- Full Root Access
- Non-Oversold Cores
- 100% NVMe Storage
- Network Firewall/IDS
- Free Daily Snapshots
- DDoS-Protected Network
- Dual Backups (Optional)
- Fully Managed (Optional)
30-day money-back
SC-8GB
- 4 vCPU
- 8 GB RAM
- 200 GB NVMe
- 20 TB transfer
- Full Root Access
- Non-Oversold Cores
- 100% NVMe Storage
- Network Firewall/IDS
- Free Daily Snapshots
- DDoS-Protected Network
- Dual Backups (Optional)
- Fully Managed (Optional)
30-day money-back
SC-16GB
- 6 vCPU
- 16 GB RAM
- 400 GB NVMe
- 40 TB transfer
- Full Root Access
- Non-Oversold Cores
- 100% NVMe Storage
- Network Firewall/IDS
- Free Daily Snapshots
- DDoS-Protected Network
- Dual Backups (Optional)
- Fully Managed (Optional)
30-day money-back
AnythingLLM installs itself on first boot with TLS, a hostname, and its database wired. Management and Dual Backups are optional addons in the cart.
Why here
Why run AnythingLLM on RemarkableCloud.
Your own private VPS
Not a container on a shared box. A real virtual server with dedicated resources and your own IP, and nobody else on it.
Stack apps, one price
Add as many apps as fit on your server from the client area, in one click each. No per-app charge, ever.
A 500% SLA that pays
1 hour down = 5 hours credited back, automatically, from the first minute. While others credit after thresholds, we credit from minute one.
Non-oversold hardware
Enterprise dedicated infrastructure, 100% NVMe, resources never sold twice. Your AnythingLLM performs the same on the worst day as the best.
Humans since 2001
Real engineers answer, at any hour. Twenty-five years of running production servers, and no chatbot wall between you and them.
Free migration, done by us
Moving an existing AnythingLLM site? We migrate it, verify it, and hand it back running, at no cost. Your old host stays live until you approve.
Built for
What people run on a AnythingLLM VPS here.
Chat with your document library
Workspaces of PDFs, docs and sites, queried by RAG that cites what it used.
A private research assistant
Internal reports and contracts become conversational without leaving your server.
Team knowledge workspaces
Per-workspace documents and permissions, so sales and engineering query their own corpus.
Agent skills over your data
Beyond Q&A: agents that browse, summarize and act, grounded in your library.
The details
AnythingLLM here, in plain words.
What can I actually do with AnythingLLM?
Feed it contracts, manuals, research, or an exported wiki and ask questions in plain language; answers cite the source passages. Teams use workspaces to separate clients or projects, each with its own documents and chat history.
Does my data stay private?
Documents and embeddings live on your server, inside our dual offsite backups. If you pair it with Ollama on the same Cube, even inference stays local and nothing ever leaves the machine; with a commercial API key, only the retrieved passages travel to the provider.
How much RAM does AnythingLLM need?
4 GB runs the app and its embedded vector database for serious document sets. Add Ollama for local inference and the pair wants 16 GB or more depending on the model; the cart meter does that math for you.
Which models work with it?
OpenAI, Anthropic, Google, and any OpenAI-compatible endpoint, plus local models through Ollama. Embeddings can run locally either way, which keeps ingestion free and private.
Everything around AnythingLLM, in one honest list.
Everything unmarked ships with every plan. The starred items are optional paid addons, priced plainly, added in the cart or anytime later.
* Optional paid addon: part of Fully Managed or Dual Backups. One price per plan, addons priced plainly, no surprise renewals.
Stack it
AnythingLLM runs well with.
Same server, no extra bill. These are the companions our team installs next to AnythingLLM most often.
Ollama
+16 GBRun open LLMs on your own server. Honest requirement: real models want real RAM, and the meter will route you to the right plan.
LiteLLM
+1 GBOne API gateway for 100+ LLM providers: unified OpenAI-format calls, spend tracking, rate limits, and failover, self-hosted.
Nextcloud
+2 GBYour own file, calendar, and office suite. The self-hosted office anchor, with backups included instead of sold separately.
How it works
Seven systems working so you do not.
Every managed server runs the same stack underneath. Pick a system to see what it actually does.
Issues end before you hear about them.
Every minute, we check more than 200 parameters of your server and its services: web server, database, mail, disk, memory, and the network path to it. The moment anything drifts out of range, our support staff is alerted, investigates, and fixes it.
This is internal monitoring, deeper than any public uptime check: it watches the inside of your server, not just whether it answers. Most incidents are found and fixed by our staff before anyone outside ever notices them.
- 200+ parameters checked every minute
- Alerts reach engineers, not ticket queues
- Inside-the-server checks, not just ping
- Humans respond in minutes, any hour
What customers say
More reviews: Trustpilot“Excellent webhosting company. I really love all their products and services. Thank you so much.”
“The tech support is excellent, and they offer very competitive pricing. They also migrate my cPanel VPS to a DirectAdmin VPS without noticeable downtime.”
“I've been hosting several websites with them, and I can honestly say their support is outstanding. They are incredibly patient, knowledgeable, and genuinely committed to solving problems, even the most difficult and unusual ones.”
Frequently asked questions
Is AnythingLLM good for a small team?
Yes: multi-user mode with per-workspace permissions is built into the Docker deployment we ship. One server holds separate workspaces per client or department cleanly.
AnythingLLM or a custom RAG build?
AnythingLLM gets you a working, private document-chat system the day the server provisions. A custom Flowise pipeline wins when you need exotic retrieval logic; plenty of teams run both, and both are $0 here.
What does it cost to run?
The app is $0; a 4 GB server starts at $13.33/mo on annual billing, management and backups optional. Local embeddings and Ollama inference add no per-token cost; commercial API models bill with their provider.
Where do embeddings and vectors live?
On the server: AnythingLLM ships an embedded vector store (with options to use external ones), so both your documents and their embeddings stay private.
Which LLMs can it use?
Your choice per workspace: commercial APIs with your keys, or local models through Ollama, including one stacked on this same server for a fully private pipeline.
Is this AnythingLLM on a VPS or on shared hosting?
A VPS. AnythingLLM runs on your own private virtual server with dedicated, non-oversold resources and its own IP. You get real isolation and performance, with the setup, TLS and hardening handled for you.
Can I run other apps on the same server?
Yes. Add any other catalog app to this server from your client area, one click each, no per-app charge. How much fits is a question of RAM, and resizing is minutes.
Do you migrate my existing AnythingLLM?
Yes, migration help is free. Open a ticket with access to the current setup and we move it, verify it, and hand it back running. Nothing switches until you approve.
Can I upgrade or downgrade the plan later?
Anytime. Plans resize from the client area in minutes, annual billing keeps its flat 17% discount at every size, and there is never a fee to change tiers.
Sizing an app stack? The VPS sizing calculator recommends a plan from your apps and traffic, with the math shown.
Shared CPU servers
The right home for AnythingLLM and most stacks: 3.0+ GHz vCPU, from $4.17/mo.
See Shared CPU →Dedicated CPU servers
For CPU-hungry stacks and busy databases: cores that are physically yours.
See Dedicated CPU →All plans and pricing
Every plan, both families, one honest price with everything included.
See pricing →Your server runs. You sleep.
Fully managed hosting from people who have been doing this since 2001.