WHERE WE DEPLOY

We optimise every layer of inference. And deploy it wherever you say.

The same layer, the same team and the same optimisations run in three different places depending on what your case needs. Starting in one and moving to another is not a migration: it is a deployment decision.

The route

The same stack, three destinations.

CLIENTSOVEREIGN AI● SPAIN · EUYOUR CLOUDON-PREMISENON-EU

On our nodes

Our own infrastructure in Spain. You manage nothing: not hardware, not stack, not scaling.

api.nextbit256.com · Running
Regioneu-south · EU
Uptime 90 d99,9 %
Node local time--:--:--
Serverless · Dedicated

In your cloud

We deploy and operate our stack inside your AWS, Azure or GCP account. You leverage your spend commitment and your network policies.

AWSAzureGCP
Runs inYour account onAWSAzureGCP
RegionThe one you choose
eu-west-1us-east-1me-central-1

Final availability depends on the hyperscaler.

Managed Dedicated Inference

On your premises

Real on-premise: physical hardware in your datacenter, zero data egress. Operated end to end by us.

Sovereign deployment
Decision table

What changes between them.

On our nodesDedicated / Capacity blockOn your premises
IsolationMulti-tenant with logical isolationGPU dedicated in hardwareFully physical
JurisdictionSpain (Nextbit)Spain (Nextbit)Yours
Who operates itNextbit, everythingNextbit, everythingNextbit operates, you host
Pricing modelPer million tokensFixed monthly feeProject + fee
The same API in all three. Anyone who starts on serverless and grows into dedicated or on-premise does not rewrite a single line of integration.