Skip to main content
SHEET 01On-Premise DeploymentSOVEREIGN

Your platform. Your data center. Your rules.

Deploy the entire assistents platform, including self-hosted LLMs, in your data center or private cloud. Every capability available in cloud is available on-premise. Zero external API calls. Full data sovereignty.

  • Air-Gapped Ready
  • Self-Hosted LLMs
  • Zero External APIs
  • Data Sovereignty
  • FedRAMP Capable
ZeroExternal API calls required
<4 WeeksFull deployment timeline
24/7Dedicated infrastructure support
SHEET 02Full Platform, On Your TermsPARITY

Every capability available in cloud is available on-premise.

The Context Engine, Governance layer, Action Engine, and all pre-built agents run identically whether deployed in the assistents cloud or your data center. No feature gaps. No compromises.

Cloud / on-premise module mapIn sync
Cloud DeploymentParityOn-Premise Deployment
Context Engine1:1Context Engine
Governance Layer1:1Governance Layer
Action Engine1:1Action Engine
Agent Builder1:1Agent Builder
Orchestration1:1Orchestration
Canvas1:1Canvas
Data Analysis1:1Data Analysis
300+ Connectors1:1300+ Connectors

1:1 feature parity no compromises 8 modules mapped

SHEET 03Self-Hosted LLMsM-01..04

Fine-tune, deploy, and run LLMs on your own hardware.

Full control over model selection, training data, and inference. Deploy open-source models like Llama, Mistral, or Qwen on your infrastructure. Fine-tune models on your proprietary data without sending anything external. All deployment logs remain within your environment.

Full control over model selection, training data, and inference. Deploy open-source models on your infrastructure.

Fine-tune on your proprietary data. All training data, model weights, and inference logs remain within your controlled environment.

Use commercial LLMs via your own API keys, or run fully open-source stacks with zero external dependencies.

Learn more
Private network · secure perimeter
PERIMETER · ON-PREMISEM-01Base ModelLlama, Mistral, QwenM-02Fine-TuningYour proprietary dataM-03Model RegistryVersion controlledM-04Inference EngineGPU cluster optimizedAgent DeploymentAll agents use your self-hosted models

No external API calls

SHEET 04Data Sovereignty & ComplianceSOV-01..03

Meet your strictest infrastructure requirements.

From data isolation to regulatory compliance to fully air-gapped environments, the assistents platform meets your most demanding infrastructure and security requirements.

SOV-01

Data Sovereignty

Zero data leaves your infrastructure. All processing, storage, and inference happens in your controlled environment. Your data never touches external infrastructure — ever.

  • Full data isolation
  • Zero egress traffic
  • AES-256 encryption at rest
  • Customer-managed keys
SOV-02

Regulatory Compliance

Meet HIPAA, SOC 2, FedRAMP, GDPR, and the strictest data residency requirements without compromises or workarounds.

  • HIPAA compliant
  • SOC 2 Type II
  • FedRAMP capable
  • GDPR data residency
SOV-03

Air-Gapped Deployment

Deploy fully disconnected from the internet. Completely isolated environments for maximum security in defense, intelligence, and critical infrastructure.

  • No internet required
  • Offline update packages
  • Local DNS resolution
  • HSM key management
SHEET 05Deployment Options5 MODELS

Healthcare, defense, and financial services teams trust assistents.

Five deployment models to match your security posture, compliance requirements, and operational preferences. From fully managed cloud to completely air-gapped on-premise installations.

DM-01

Cloud

Multi-tenant SaaS. Fast onboarding, fully managed, global availability.

DM-02

Private Cloud

Single-tenant in AWS, Azure, or GCP. Isolated infrastructure.

DM-03Primary

On-Premise

Full platform in your data center. Complete control, data residency, network isolation.

DM-04

Hybrid

Mix cloud and on-premise. Cloud analytics with on-prem agents.

DM-05

VPC

Dedicated VPC in your AWS or Azure account. Managed by assistents.

Cloud / hybrid / on-premise topology
NODE-01Cloud Gatewayassistents-managed cloudNODE-02Hybrid BridgeSelective sync layerNODE-03On-Prem Data CenterFull stack isolated
5Deployment options
ZeroExternal API calls
<4 WeeksFull deployment
99.99%Uptime SLA
SHEET 06Deployment DashboardCONTROL PLANE

Infrastructure monitoring and health at a glance.

Monitor all infrastructure from a single pane of glass. GPU utilization, model health, query latency, and service availability, with real-time alerts for anomalies and automated incident escalation.

Monitor all infrastructure from a single pane of glass. GPU utilization, model health, query latency, and service availability.

Real-time alerts for anomalies. Automated incident escalation. Historical performance trending and capacity planning built in.

Integrates with your existing monitoring stack: Datadog, Grafana, PagerDuty, and more.

Explore dashboard
Sys statusOperationalUptime: 99.99%
GPU Utilization78%8x A100 cluster
Memory Usage64 GBof 128 GB allocated
Active Models123 fine-tuning in progress
Inference Latency45msp99 across all models
Service health
Inference Engineonline99.99%
Model Registryonline99.98%
Fine-Tuning Pipelineonline99.97%
Data Connector Hubonline99.99%
Governance Engineonline99.99%
SHEET 07Sign-offREADY

Meet the strictest data residency requirements.

Healthcare, defense, and financial services teams run assistents in air-gapped environments. Deploy the full platform in your data center with dedicated infrastructure support.

Workload
Healthcare · defense · financial services
Egress
Zero external API calls · air-gapped
Lead time
Full platform deployment · 4 weeks
Sheet
07 of 07 · Model Hub

Infrastructure assessment

Evaluate your environment, compliance requirements, and deployment options with our infrastructure specialists.

Book a consultation

Review deployment guide

Architecture patterns, scaling strategies, and best practices for running assistents on-premise across organizations.

How It Works