singletenant.ai

Vendors · Comparison

Scaleway vs Together AI

Side-by-side comparison of Scaleway and Together AI for single-tenant LLM hosting. Deployment options, compliance, pricing, and operational fit compared.

Scaleway

Managed Inference runs models on dedicated GPUs at fixed hourly rates (e.g. Llama 3.1-8b at 0.93 EUR/hour) inside a VPC; prompts and responses are stated as stored only in Europe. Scaleway SAS is registered in Paris, France (legal notice). Security page says the SecNumCloud qualification process was entered in January 2025; not yet listed as qualified.

Together AI

Dedicated endpoints are explicitly single-tenant GPU instances with public per-hour pricing. Terms of service make no geographic storage commitment, so residency guarantee is recorded as none.

From the data

Key differences

  • Category Scaleway is a EU sovereign offering; Together AI is Dedicated GPU.
  • Compliance Only Scaleway documents ISO 27001 and GDPR DPA. Only Together AI documents SOC 2 Type II and HIPAA BAA.
  • Data residency Scaleway has a verified residency guarantee; Together AI offers no residency guarantee (verified).
  • Jurisdiction Parent jurisdiction: Scaleway FR, Together AI US.
  • Pricing Scaleway: Per-hour dedicated GPU (Managed Inference); per token (Generative APIs) (published pricing). Together AI: Per 1M tokens serverless; per GPU-hour for dedicated endpoints and clusters (published pricing).

Generated from the verified vendor data below; sources and dates on each vendor profile.

Side by side

Capabilities compared

Scaleway Together AI
Category EU sovereign Dedicated GPU
Deployment models Serverless APIDedicated GPU inference ServerlessDedicated endpointGPU clusters
Regions Europe Not verified
Compliance ISO 27001GDPR DPA SOC 2 Type IIHIPAA BAA
Pricing model Per-hour dedicated GPU (Managed Inference); per token (Generative APIs) Per 1M tokens serverless; per GPU-hour for dedicated endpoints and clusters
Public pricing Yes Yes
Residency guarantee Yes No
Parent jurisdiction FR US
Analyst note Managed Inference runs models on dedicated GPUs at fixed hourly rates (e.g. Llama 3.1-8b at 0.93 EUR/hour) inside a VPC; prompts and responses are stated as stored only in Europe. Scaleway SAS is registered in Paris, France (legal notice). Security page says the SecNumCloud qualification process was entered in January 2025; not yet listed as qualified. Dedicated endpoints are explicitly single-tenant GPU instances with public per-hour pricing. Terms of service make no geographic storage commitment, so residency guarantee is recorded as none.

Where they diverge

Deployment differentiation

Only Scaleway

Serverless APIDedicated GPU inference

Both

No overlap in deployment models.

Only Together AI

ServerlessDedicated endpointGPU clusters

Read the full profiles