Enterprise GPU Cloud.
Built in the Netherlands.

Dedicated NVIDIA GPU capacity for AI training and inference, delivered with EU data residency and direct hardware ownership.

EU data residency Dedicated, owned GPU hardware Direct AMS-IX connectivity Netherlands registered
New from NovaServe

The NovaServe Token Factory

The industry is shifting how it measures AI infrastructure. Instead of GPU-hours, the unit that matters is the token: every word generated, every inference step, every agent action. NVIDIA has termed this shift the "AI factory" — compute built to produce tokens at scale, the way a power plant produces electricity.

The NovaServe Token Factory puts a metered, pay-per-token API layer on top of our own GPU fleet. You get inference priced by output, not by idle capacity — running on hardware we own, hosted in the Netherlands.

€/1M tokens
Priced by output, not by the hour
EU-only
Hosted entirely in the Netherlands
Open models
Llama, Mistral, Qwen & more
Dedicated
Reserved throughput, no queueing

Metered by the token

Pay per million tokens processed, not per GPU-hour reserved. Costs track usage directly.

Hosted in the Netherlands

Inference runs exclusively on our own Dutch infrastructure, not shared with third-party tenants.

Open-weight models

Serve leading open models out of the box, or bring your own fine-tuned weights.

Dedicated throughput

Reserved-capacity tiers remove multi-tenant queueing for latency-sensitive workloads.

Standard
from €0.35per million tokens
Shared token pool. Ideal for prototyping and variable workloads.
Dedicated
from €2.10per hour, reserved
Reserved throughput on H200 capacity for production traffic.
Enterprise
CustomSLA & volume pricing
Committed capacity, custom models, and a dedicated support line.

Indicative pricing. Final rates depend on model, context length and committed volume.

GPU capacity

Pricing by GPU tier

Hourly on-demand pricing across our current fleet. All GPUs are dedicated, not shared or virtualised.

Specification RTX 4090 RTX 5090 RTX Pro 6000 H200
Flagship
VRAM 24 GB GDDR6X 32 GB GDDR7 96 GB GDDR7 ECC 141 GB HBM3e
Memory bandwidth 1.0 TB/s 1.79 TB/s 1.6 TB/s 4.8 TB/s
FP16 performance ~330 TFLOPS ~419 TFLOPS ~503 TFLOPS ~989 TFLOPS
FP8 performance ~660 TFLOPS ~838 TFLOPS ~1,006 TFLOPS ~1,979 TFLOPS
Best for Rendering, dev & test Generative media, fine-tuning Inference at scale LLM training & large-context inference
Indicative price from €0.65/hron-demand from €1.05/hron-demand from €1.85/hron-demand from €3.40/hron-demand

Prices shown are indicative on-demand rates per GPU. Reserved and enterprise pricing, with volume discounts, is available on request.

Applications

Matched to the workload

We size the hardware to the job, not the other way round.

H200

LLM training & fine-tuning

Large-context model training and fine-tuning workloads that need maximum memory bandwidth.

RTX Pro 6000 / 5090

Inference at scale

High-throughput serving for production inference endpoints with predictable latency.

RTX 4090 / 5090

Rendering & generative media

Image, video and 3D rendering pipelines that benefit from strong single-GPU throughput.

Any tier

Research & experimentation

Short-lived, flexible capacity for prototyping models and testing new architectures.

Why NovaServe

Infrastructure you can audit

Built for organisations that need to know exactly where their data and compute sit.

European sovereignty

Data stays in the EU

All compute and storage remain within the Netherlands and the EU, with no dependency on non-EU cloud infrastructure.

Direct ownership

Hardware we own and run

We operate our own GPU fleet. No reseller layer, no capacity shared with third-party tenants.

Service

White-glove support

A dedicated technical team for onboarding, scaling and incident response, from first call to production.

Get in touch

Talk to our team

Tell us about your workload and we'll come back with capacity and pricing options within one business day.

Location

Amsterdam, Netherlands

Enterprise & reserved capacity

Available on request
Please enter your name.
Please enter your company.
Please enter a valid email address.
Please select an option.
Please tell us a little about your workload.

Message sent

Thank you. A member of our team will be in touch within one business day.