AI SERVERS — COMPARE
Every Zanus customer sits somewhere on the ownership ladder: a hosted tenant (start in days), the On-Prem Tokens Server (your tokens, your GPUs — the Quantum without a tenant), or the Private On-Premises AI (the whole working system with your industry tenant, owned). Same software at every rung — you climb when the economics or the regulations say so.
Same UX at every rung · your knowledge and configuration move with you

One machine family, two products — sized on your models, context and tokens per day.
| Hosted tenant (SaaS) | On-Prem Tokens Server | Private On-Premises AI | |
|---|---|---|---|
| What you get | Front Office / Back Office in the Zanus datacenter | The Quantum + Zanus OS — pure token endpoint, no tenant | Server + full Zanus AI OS + your industry tenant(s) included |
| Where your data lives | Zanus datacenter — isolated tenant | Your building — prompts never leave | Your building — air-gap capable, never on the internet |
| What you pay | Flat yearly plan + credits | One purchase (RFQ) + electricity | One purchase + permanent app licenses (RFQ) — no recurring fees or tokens |
| AI usage cost | Credits — 1 credit = 1¢ | Cents of electricity after payback | Cents of electricity — nothing metered, ever |
| Live in | Days | ~3 weeks, delivered configured | ~3 weeks, delivered configured |
| Best when | You want to start now and prove value | High volume — hundreds of thousands of interactions a month | Data that can never touch the internet; zero recurring fees |
No — it is the same OS and the same UX at every rung. Your knowledge, configuration and history move with you.
Yes. Hosted tenant + your own token server is the most popular hybrid: SaaS convenience, on-prem economics.
Because honest sizing depends on the models you want, the context you need and your tokens per day. You tell us the workload; we quote the machine — no oversized invoice, no undersized server.
← Previous: Private On-Premises AI · Sharing with a colleague? 📄 Get the PDF · ✉️ Email this page ·