On-premise, air-gapped and private-cloud AI

Enterprise AI that never leaves your building.

We deploy open-source language models on your own hardware or private cloud — sized, quantized, fine-tuned on your data and handed over with full documentation. No per-token fees. No data leaving your perimeter.

Why private AI

Cloud AI is convenient. It is not always allowed.

Regulated data, client privilege, trade secrets and sovereignty rules stop many organisations from sending prompts to third-party APIs. Open-weight models now match cloud quality on most enterprise tasks — the missing piece is an engineering partner to deploy them properly.

Public cloud AI Exposure

  • Every prompt and document crosses your boundary
  • Per-token fees grow with every successful rollout
  • Vendor price changes and deprecations are your risk
  • Compliance must defend an external processing relationship

2oo.one private AI Contained

  • Inference, retrieval and weights stay on your hardware
  • Fixed infrastructure cost - usage growth is free
  • You own the stack; no provider can switch it off
  • Audit logs land in your SIEM, not a vendor dashboard

Data stays inside

Prompts, documents and model weights never cross your network boundary. Optional full air-gap for classified or high-sensitivity environments.

Fixed, predictable cost

Above a few million tokens a day, self-hosting typically beats API pricing. Hardware is a one-time cost you depreciate — not a per-seat subscription.

No vendor lock-in

You own the stack. A provider price change, outage or product shutdown cannot take your AI capability offline.

Services

One partner from hardware sizing to handover

On-premise AI deployment

GPU sizing, model selection, inference serving, SSO, metering and quota controls — installed in your data centre or server room.

Explore the service

LLM fine-tuning

LoRA, QLoRA and full fine-tunes on your proprietary data, evaluated against your real tasks before anything ships.

Explore the service

Air-gapped AI

Fully disconnected deployments for defence, government and IP-critical R&D, including offline model update procedures.

Explore the service

Cloud-to-on-prem migration

Move workloads off OpenAI, Anthropic or Azure endpoints onto self-hosted models with API-compatible drop-in replacements.

Explore the service

Managed private AI

Monitoring, model refreshes, security patching and capacity planning after go-live — your AI stack, our pager.

Explore the service

20+ proven use cases

Unified search, support agents, document intelligence, email and call triage, surveillance analytics and more — with the model types behind each.

Browse use cases
How we work

A deployment process built for auditors as much as engineers

  1. Assess

    Workload analysis, data classification, compliance constraints and a hardware sizing report with transparent cost modelling.

  2. Pilot

    A working proof of concept on representative data — quantized open models served on rented or existing hardware, measured against your acceptance criteria.

  3. Deploy

    Production installation: inference servers, gateway, access control, logging and integration with your identity provider and internal tools.

  4. Tune

    Fine-tuning and retrieval pipelines built on your documents, with evaluation suites you can re-run after every model update.

  5. Hand over or manage

    Full documentation and training for your team — or a managed service agreement if you would rather we keep operating it.

Deployment, automated

From bare metal to serving in one pipeline

Our provisioning tooling takes a GPU node from empty to a governed, OpenAI-compatible endpoint — model pull, quantization, serving, gateway, SSO and metering — reproducibly, and documented so your team can re-run it without us.

provision - onprem-dc1
$ 2oo provision --site onprem-dc1OK GPU node detected: 4x H100 80GB - CUDA 12.8$ 2oo deploy --model llama-4-70b --quant awq --serve vllm Pulling weights .......... done (38.2 GB, verified sha256) Quantization: AWQ int4 - eval delta vs FP16: -0.6%OK vLLM serving at https://ai.internal:8000/v1$ 2oo gateway --sso oidc --meter per-team --quota finance=2M/dayOK Gateway live - SSO connected - metering on - audit to SIEM$ 2oo verify --external-callsOK 0 external endpoints reached - perimeter sealed
Platform capabilities

Enterprise controls, built in from day one

A model server alone is not a deployment. Every system we ship includes the governance layer your IT and finance teams will ask about in week one.

Single sign-on

SAML and OIDC integration with Entra ID, Okta, Keycloak or your existing identity provider — role-based access mapped to your directory groups.

Usage metering & chargeback

Per-user, per-team and per-application token metering with exportable reports, so departments can be billed internally and ROI is measurable.

Quota & budget management

Hard and soft limits per key, team or project; rate limiting, priority tiers and alerts before anyone exhausts shared capacity.

Advanced retrieval: RAG & GraphRAG

Vector search (Qdrant, Milvus, pgvector) plus knowledge-graph retrieval for questions that span entities and relationships across your documents.

Observability & evaluation

Tracing, latency and quality dashboards with self-hosted tooling, and evaluation suites you can re-run after every model or prompt change.

Audit & guardrails

Complete prompt/response audit logs to your SIEM, PII redaction filters and policy-based routing for sensitive content.

Compliance & governance

Compliance is the architecture, not a checkbox

Frameworks are easiest to satisfy when the data simply never leaves. We design every deployment so your auditors review controls you own — and we deliver the architecture, data-flow and control documentation their reviews require. We support your certification programme; certifications themselves remain yours to hold.

GDPR
Data residency by design
HIPAA
PHI never transits out
SOC 2
Control-mapped architecture
ISO 27001
ISMS-ready documentation
EU AI Act
Deployment-risk documentation
CMMC / ITAR
Air-gap capable
Industries

Built for organisations that cannot use public AI

Healthcare

Clinical documentation and PHI-safe assistants aligned with HIPAA.

Legal

Document review and research without breaking client privilege.

Financial services

Analysis and automation inside your regulatory perimeter.

Government & defence

Air-gapped and sovereignty-compliant deployments.

Insurance & actuarial

Claims, underwriting and actuarial document automation.

Banking, identity & fraud

KYC document intelligence and fraud-investigation support.

Customer experience

Support deflection, agent assist and QA at fixed cost.

All industries

See every sector we deploy for, with compliance notes.

Ready to bring AI inside your perimeter?

Book a consultation