ChatVLM Enterprise

Frontier AI, from private cloud to air gap

Every ChatVLM plan carries contractual zero data retention and no training on customer data. Enterprise goes further, running open-weight frontier models on private infrastructure or entirely on-site.

Private models without the hardware

The same Pro and Ultra seat tiers as Teams, plus the identity and governance controls a security review asks for.
Enterprise
Private cloud
Pro and Ultra seats for the whole organization, plus private models, identity controls, and compliance infrastructure.
  • Everything in Teams
  • Pro & Ultra seat tiers
  • Sovereign Presets on private infra
  • Zero third-party subprocessors
  • SAML SSO + SCIM provisioning
  • Workspace model allowlisting
  • Custom retention & legal holds
  • Immutable audit logs & SIEM export
  • SOC 2 Type II & HIPAA BAA package
  • Dedicated CSM & security review help
Custom
Contact sales
25-seat minimum · Annual, invoiced billing

Sovereign Box: the whole stack in one rack

Two appliances, one sized for an office and one for a datacenter. Both ship as clusters of at least three nodes, with unlimited seats, offline license enforcement, and quarterly model updates on encrypted physical media.
Sovereign Box Theta
The office-ready private AI server. High-concurrency local routing on standard wall power.
From $85,000
per node, one-time hardware+ $60,000/yr cluster license
  • 4× RTX PRO 6000 Blackwell Server Edition, 384 GB VRAM
  • Multiple mid-size open models, routed intelligently
  • Up to 4 models served simultaneously
  • From 150 concurrent streams and 1,000 users
  • +50 streams per node beyond the third
  • Full capacity with two nodes down
  • Standard 110/220 V office power
  • Self-contained, with database, cache, and storage on board
  • Unlimited seats
  • Quarterly offline model updates
Talk to sales
Priced by CPU/RAM configuration · Or from $12,200/mo as HaaS for a three-node cluster (36-month term)
Sovereign Box Omega
Air-gap flagship
The datacenter deployment, for organizations standardizing on the largest open-weight models.
$580,000
per node, one-time hardware+ $150,000/yr cluster license + Sovereign Control Plane: $110,000 hardware + $20,000/yr Or bring existing infrastructure to our reference architecture
  • 8× NVIDIA HGX B300, 2.3 TB HBM3e
  • Largest open-weight models resident in memory
  • Headroom for smaller routing models
  • From 600 concurrent streams and 4,000 users
  • +200 streams per node beyond the third
  • Full capacity with two nodes down
  • 24/7 monitoring, 4-hour P1 response
  • Unlimited seats
  • Secure courier-delivered updates
Talk to sales
Or from $65,500/mo as HaaS for a three-node cluster with control plane · Pre-sale site survey required

One private endpoint for every AI tool

Enterprise endpoints and Sovereign Boxes speak the standard APIs these tools already use, so coding agents, IDE assistants, and internal automations all run on private models.
  • Coding agents on private models
    Point Claude Code or any other AI assistant at the private endpoint. Source code, prompts, and diffs stay on infrastructure the company controls.
  • Standard, drop-in APIs
    Industry-standard API compatibility means existing SDKs, IDE plugins, and internal tools work unchanged. Swap the base URL, keep the workflow.
  • Internal automations, same perimeter
    Nightly extraction jobs, CI reviews, and in-house agents call the same endpoint, so prompts, source code, and documents never leave the deployment.

As many agents as you want, for the same bill

Teams arrive with an agentic workflow whose API bill grows every month it works well, and leave running the same workflow flat out on machines they own. On a Sovereign Box the marginal token is free; a private cloud is priced on capacity, not usage.
  • The meter is off
    You pay for capacity: a box in your rack, or single-tenant infrastructure we operate. Whether it idles or runs all night, the invoice at the end of the month is the one you signed.
  • Agents without a quota
    Coding agents, extraction pipelines, CI reviews, and overnight evaluations run at once, around the clock. The tenth agent costs what the first one did.
  • More tokens for the same money
    Nobody trims context to save cents or thinks twice about a retry. The work gets the tokens it actually needs, and the number on the invoice stays where it was.

When the big providers go down, you don't

A public API is one shared system having one shared bad day. A Sovereign Box serves your people and nobody else's, from hardware in your own building, so an incident somewhere else is not an outage for your company.
  • Nobody else's traffic, nobody else's incident
    The cluster serves your users and no one else, so another company's spike or a provider's control-plane failure has no path to it. An air-gapped Box does not even need the internet to be up.
  • Provider outages pass you by
    Large providers go down for hours at a time. Work on a Box has carried on straight through those incidents, because nothing in the answer path belongs to a company you do not control.
  • Full capacity with two nodes down
    Every cluster ships as at least three nodes, so a failed node, or one pulled for maintenance, does not take the service with it. Updates land in your change window, not when a vendor ships.

Total cost of ownership

Metered seats and agentic usage scale with every person added. An appliance doesn't. Set the headcount below to see where the lines cross.

Marketing, sales & operations

Drafting, research, summarizing, everyday chat

Analysts, legal & research

Long documents, deep research, council answers

Engineers

Coding assistants and agentic workloads

How hard do your engineers push it?

Agentic workloads are the line that dominates a metered bill.

Sovereign Box support

Per managed node equivalent, with the first three included across the fleet.

Saved per month

$12,800

Saved per year

$153,600

Running 100 people on your own hardware, against the same org on metered Enterprise SaaS seats.

Enterprise SaaS

$25,000/mo

Seats
$5,000
Agentic usage
$20,000

3 × Sovereign Box Theta

$12,200/mo

3-node HA cluster, automatic failover
3 nodes (hardware)
$7,200
Software license (cluster)
$5,000
Standard support
Included
Seats
Unlimited
Agentic usage
$0 (marginal token is free)

Serving 100 people and ~30 agents

22 of 150 streams

Sizing is set by how many people generate at the same instant rather than by headcount, because most of an org is reading rather than waiting on tokens. At this mix the configuration has room for about 560 people before you need more capacity.

The bar shows whichever ceiling sets the node count, streams or population. Committed capacity here is 150 streams across 3 nodes at 50 each, with 1,000 rated seats. Each node commits a third of what it can physically serve, so 2 nodes' worth stays in reserve and losing any two of them serves this load without interruption.

Estimates only, for comparison. Seats priced at the published Pro ($30) and Ultra ($80) rates; appliances at Hardware-as-a-Service monthly pricing on a 36-month term, software license included. Agentic figures assume $1,000/engineer/month of metered model usage at standard intensity, and about 1.5 concurrent agents per engineer. Clusters start at 3 nodes, hold 2 nodes' worth of capacity in reserve, and are sized to stay under 80% of committed streams; every node bills the same monthly rate under a single cluster-wide software license. A Theta cluster scales horizontally until metered spend passes $131,000/month, where Omega and the largest open-weight models become the better step. Support is shown at the tier selected above, priced per managed node equivalent with the first 3 MNE included once across the fleet.

Frequently asked questions

Add-ons and services

Listed up front rather than averaged into the quote.

  • Additional sites$25,000/yr beyond the first
  • Three-year prepayment10% discount
  • Hardware refresh20% trade-in credit at year three
  • Pre-sale site survey$25,000, credited against purchase
  • Sovereign model evaluation on in-house data$25,000, credited against purchase
  • Theta installation and commissioningFrom $10,000
  • Omega installation and commissioningFrom $50,000
  • On-site spares cache$50,000
  • RAG and data-source integrationFrom $50,000
  • Domain distillation and model tuningFrom $75,000

Put frontier AI where the sensitive work happens

Talk to us about Sovereign Presets, a proof of concept on a Sovereign Box, or a compliance review.