singletenant.ai

Vendors · Comparison

Azure OpenAI Provisioned Throughput (PTU) vs Scaleway

Side-by-side comparison of Azure OpenAI Provisioned Throughput (PTU) and Scaleway for single-tenant LLM hosting. Deployment options, compliance, pricing, and operational fit compared.

Azure OpenAI Provisioned Throughput (PTU)

Provisioned throughput is now a Microsoft Foundry (formerly Azure AI Foundry) deployment type billed hourly per PTU, with 1-month or 1-year Azure Reservations. Global, Data Zone (US/EU), and Regional options set where prompts are processed; stored data stays in the customer-designated geography. FedRAMP scope names Azure OpenAI; SOC 2, ISO 27001, and C5 attest the Azure platform.

Scaleway

Managed Inference runs models on dedicated GPUs at fixed hourly rates (e.g. Llama 3.1-8b at 0.93 EUR/hour) inside a VPC; prompts and responses are stated as stored only in Europe. Scaleway SAS is registered in Paris, France (legal notice). Security page says the SecNumCloud qualification process was entered in January 2025; not yet listed as qualified.

From the data

Key differences

  • Category Azure OpenAI Provisioned Throughput (PTU) is a Hyperscaler dedicated offering; Scaleway is EU sovereign.
  • Compliance Both document ISO 27001 and GDPR DPA. Only Azure OpenAI Provisioned Throughput (PTU) documents SOC 2 Type II, C5, FedRAMP and HIPAA BAA.
  • Data residency Both have a verified data residency guarantee.
  • Jurisdiction Parent jurisdiction: Azure OpenAI Provisioned Throughput (PTU) US, Scaleway FR.
  • Pricing Azure OpenAI Provisioned Throughput (PTU): Per PTU per hour; Azure Reservations discount with 1-month or 1-year terms (published pricing). Scaleway: Per-hour dedicated GPU (Managed Inference); per token (Generative APIs) (published pricing).

Generated from the verified vendor data below; sources and dates on each vendor profile.

Side by side

Capabilities compared

Azure OpenAI Provisioned Throughput (PTU) Scaleway
Category Hyperscaler dedicated EU sovereign
Deployment models Global provisionedData-zone provisionedRegional provisioned Serverless APIDedicated GPU inference
Regions Global, US, Europe Europe
Compliance SOC 2 Type IIISO 27001C5FedRAMPHIPAA BAAGDPR DPA ISO 27001GDPR DPA
Pricing model Per PTU per hour; Azure Reservations discount with 1-month or 1-year terms Per-hour dedicated GPU (Managed Inference); per token (Generative APIs)
Public pricing Yes Yes
Residency guarantee Yes Yes
Parent jurisdiction US FR
Analyst note Provisioned throughput is now a Microsoft Foundry (formerly Azure AI Foundry) deployment type billed hourly per PTU, with 1-month or 1-year Azure Reservations. Global, Data Zone (US/EU), and Regional options set where prompts are processed; stored data stays in the customer-designated geography. FedRAMP scope names Azure OpenAI; SOC 2, ISO 27001, and C5 attest the Azure platform. Managed Inference runs models on dedicated GPUs at fixed hourly rates (e.g. Llama 3.1-8b at 0.93 EUR/hour) inside a VPC; prompts and responses are stated as stored only in Europe. Scaleway SAS is registered in Paris, France (legal notice). Security page says the SecNumCloud qualification process was entered in January 2025; not yet listed as qualified.

Where they diverge

Deployment differentiation

Only Azure OpenAI Provisioned Throughput (PTU)

Global provisionedData-zone provisionedRegional provisioned

Both

No overlap in deployment models.

Only Scaleway

Serverless APIDedicated GPU inference

Read the full profiles