Skip to main content

Choosing the Right Plan

Managed CaseDesk is a flat monthly subscription. Choose the plan that matches your model size and concurrency requirements.

Plans

PlanMonthly priceModel sizeRuntimeMax concurrent users
Starter£199 / €229 / $2391–8BOllama (sequential)1
Team£449 / €499 / $5199–20BvLLM20
Advanced£1,199 / €1,349 / $1,39921–70BvLLM50
Enterprise IsolationCustomAnyDedicated hardware nodeCustom

Prices are per deployment, billed monthly. No GPU-hour charges. No per-seat charges.

Contact sales for Enterprise Isolation pricing.

Concurrency

Starter uses Ollama, which processes requests sequentially. It is suited to single-user or low-concurrency workloads.

Team and Advanced use vLLM continuous batching. Multiple users can generate responses simultaneously from a single pod:

  • Team — up to 20 concurrent users generating at once
  • Advanced — up to 50 concurrent users generating at once

The flat monthly price covers all concurrent users with no additional per-seat charges.

Choosing a model

Within each plan, you choose a model from the catalogue at deployment time. Guidance by use case:

Use caseRecommended model range
Assistants, chatbots, customer support1–8B (Starter)
Complex reasoning, agents, document analysis9–20B (Team)
Large-context tasks, long documents, legal/medical analysis21–70B (Advanced)

:::tip Evaluating hosted alternatives? If you're comparing CaseDesk against GPU cloud or serverless inference providers, see the comparison pages for a side-by-side breakdown on cost, data ownership, and control. :::