Inference and RAG
Six NVIDIA GPUs in Switzerland. Datasets and secrets in-country.
Hikube for AI provides six dedicated NVIDIA GPUs (L4, L40S, A100-80, RTX 6000 Pro, H100, H200) on VMs or Kubernetes, with object storage and Vault in the same Swiss jurisdiction: for models and data that must not leave the country. All six cards are in the 14-day trial.
- 1. Corpus on S3
- 2. Keys in Vault
- 3. GPU attached
- ISO 27001
- 3 DC
- SLA 99.99%
For teams that train or infer in Switzerland: the same six cards on an isolated VM or a Kubernetes node group. GPU Operator installs drivers and the device plugin. HAMi, optional, shares one card across pods and requires GPU Operator. Not CNCF AI Conformance. Catalogue: the GPU catalog. 14-day trial.
Path
01
1. Corpus on S3
Documents and checkpoints in a Swiss bucket. Object Lock if the corpus must stay immutable, encryption per bucket.
02
2. Keys in Vault
Model keys, database access and API tokens in Vault as a Service, not in a manifest. Images in the private registry.
03
3. GPU attached
Same catalogue. On K8s: GPU Operator, optional HAMi. Workers spread automatically, no DC choice.
04
API
Provisioning via API and console. Terraform: in preparation.
In detail
All six catalogue cards: L4 24G (165 CHF/month), L40S 48G (760), A100-80 (1495), RTX 6000 Pro 96G (1550), H100 80G (1999), H200 141G (2499). Published production prices. The 14-day trial, no credit card, includes them. An engineer sizes the model, duration and VPC isolation. Three Swiss DCs, Hidora SA. HA spare-capacity restart in normal mode excludes GPUs.
VM for an isolated data scientist or a long job. Kubernetes for sharing, autoscaling and addons. Same card catalogue, same jurisdiction. On the managed cluster: GPU Operator (optional addon) installs NVIDIA drivers and the device plugin. HAMi, also optional, virtualises one card across pods and requires GPU Operator. Hikube spreads workers across Geneva, Gland and Lucerne: you do not place nodes DC by DC. Platform detail: the Kubernetes platform.
S3 object storage in Switzerland. Access secrets in Vault. Images in Registry as a Service, not GHCR or ECR by default. A us-east-1 checkpoint breaks job sovereignty. Products: S3 object storage, the secrets vault, the image registry.
Hidora SA operates the DCs, GPU attachment, the Kubernetes control plane if you use it, and FR/EN support. You keep images, checkpoints, training code, pod topology spread and the VM vs node-group choice. Terraform integration in preparation: today, Hikube API and console. Not CNCF AI Conformance.
Inference in Switzerland on data that must not leave. Fine-tuning on internal datasets. Batch jobs on VMs. Card sharing via HAMi on Kubernetes. This is not a worldwide catalogue of thousands of GPUs, and not an opaque ML PaaS.
Six published SKUs, not a farm. L4 24G (165 CHF/month): inference and vision on small models. L40S 48G (760): larger inference, rendering. A100-80 (1495): fine-tuning and training. RTX 6000 Pro 96G (1550): more VRAM, workstation profile. H100 80G (1999): heavier training. H200 141G (2499): more memory for long contexts. We do not publish tokens/s or IOPS. The engineer sizes during the trial. Catalogue: the GPU catalog.
Checkpoints, weights and cold sets: S3 object storage in Switzerland. Access secrets in Vault, not a US KMS. For scratch disk on the VM or pod: network NVMe volumes, always geo-redundant across Geneva, Gland and Lucerne, synchronous (real time) or asynchronous (improved performance) replication. Sync and async do not share the same RPO; we publish neither a number nor IOPS. No unreplicated local block volume. Products: S3 object storage, block storage, the secrets vault.
No. Hikube automatically distributes workers, including GPU node groups, across Geneva, Gland and Lucerne. You do not choose the DC. Topology spread and PodDisruptionBudget for pods stay yours. HA spare-capacity restart in normal mode excludes GPUs: a card is not an interchangeable CPU worker. Platform detail: the Kubernetes platform.
No. Terraform and Cluster API integration in preparation. Today: Hikube API and console. kubectl, Helm and S3 clients stay. A tender that requires Terraform in production today does not yet match the catalogue. An engineer says so when framing the trial.
If the model or dataset touches Swiss health data, the frame is nFADP, not HDS. Hikube is not an HDS host and is not HDS-certified. Datasets, checkpoints and backups stay in Geneva, Gland and Lucerne. A US bucket breaks the file. Healthcare page: healthcare. Engineer support FR/EN.
Related products
GPU cloudSix dedicated NVIDIA cards, 24 to 141 GB.
Managed KubernetesDedicated managed Kubernetes in Switzerland, workers in your tenant.
Swiss S3S3 object storage in Switzerland: WORM, optional encryption, Swiss residency.- Vault as a ServiceSecrets and keys in Switzerland, for apps, CI and Kubernetes.
Frequently asked questions
L4, L40S, A100-80, RTX 6000 Pro, H100, H200. Published CHF/month prices. Hourly or monthly billing. All six are in the 14-day trial.
No. They are in the 14-day trial.
Yes, as optional addons on managed Kubernetes. HAMi requires GPU Operator. Detail: the Kubernetes platform.
Terraform and Cluster API integration in preparation. Today: Hikube API and console.
No. Hikube automatically distributes workers across Geneva, Gland and Lucerne. Topology spread and PDBs for pods stay on your side.
No. Hikube Managed Kubernetes is Certified Kubernetes Hosted. Hidora SA is a KCSP. We do not claim CNCF AI Conformance.
Engineer support is in French and English, from Geneva. German is not a support language.
Ready to run on 100% Swiss infrastructure?
14-day trial, no credit card. GPUs included.
