singletenant.ai

Vendors · Comparison

Azure OpenAI Provisioned Throughput (PTU) vs Vertex AI Provisioned Throughput

Side-by-side comparison of Azure OpenAI Provisioned Throughput (PTU) and Vertex AI Provisioned Throughput for single-tenant LLM hosting. Deployment options, compliance, pricing, and operational fit compared.

Azure OpenAI Provisioned Throughput (PTU)

Provisioned throughput is now a Microsoft Foundry (formerly Azure AI Foundry) deployment type billed hourly per PTU, with 1-month or 1-year Azure Reservations. Global, Data Zone (US/EU), and Regional options set where prompts are processed; stored data stays in the customer-designated geography. FedRAMP scope names Azure OpenAI; SOC 2, ISO 27001, and C5 attest the Azure platform.

Vertex AI Provisioned Throughput

Provisioned Throughput reserves Gemini and partner-model capacity in generative AI scale units (GSUs) on fixed-cost terms, including 1-week options for select models. Google Cloud docs now brand the platform Gemini Enterprise Agent Platform. Vertex AI Platform appears in Google's SOC 1/2/3 and ISO 27001 scope; public GSU price list not verified at review.

From the data

Key differences

  • Category Both are Hyperscaler dedicated offerings.
  • Compliance Both document SOC 2 Type II, ISO 27001 and GDPR DPA. Only Azure OpenAI Provisioned Throughput (PTU) documents C5, FedRAMP and HIPAA BAA. Only Vertex AI Provisioned Throughput documents SOC 3.
  • Data residency Azure OpenAI Provisioned Throughput (PTU) has a verified residency guarantee; Vertex AI Provisioned Throughput's residency guarantee is not verified.
  • Jurisdiction Both parent companies are under US jurisdiction.
  • Pricing Azure OpenAI Provisioned Throughput (PTU): Per PTU per hour; Azure Reservations discount with 1-month or 1-year terms (published pricing). Vertex AI Provisioned Throughput: Per GSU (generative AI scale unit) fixed-cost term; 1-week terms for select models (publication not verified).

Generated from the verified vendor data below; sources and dates on each vendor profile.

Side by side

Capabilities compared

Azure OpenAI Provisioned Throughput (PTU) Vertex AI Provisioned Throughput
Category Hyperscaler dedicated Hyperscaler dedicated
Deployment models Global provisionedData-zone provisionedRegional provisioned Provisioned throughputSingle-zone provisioned throughput
Regions Global, US, Europe US, Europe, Asia-Pacific
Compliance SOC 2 Type IIISO 27001C5FedRAMPHIPAA BAAGDPR DPA SOC 2 Type IISOC 3ISO 27001GDPR DPA
Pricing model Per PTU per hour; Azure Reservations discount with 1-month or 1-year terms Per GSU (generative AI scale unit) fixed-cost term; 1-week terms for select models
Public pricing Yes Not verified
Residency guarantee Yes Not verified
Parent jurisdiction US US
Analyst note Provisioned throughput is now a Microsoft Foundry (formerly Azure AI Foundry) deployment type billed hourly per PTU, with 1-month or 1-year Azure Reservations. Global, Data Zone (US/EU), and Regional options set where prompts are processed; stored data stays in the customer-designated geography. FedRAMP scope names Azure OpenAI; SOC 2, ISO 27001, and C5 attest the Azure platform. Provisioned Throughput reserves Gemini and partner-model capacity in generative AI scale units (GSUs) on fixed-cost terms, including 1-week options for select models. Google Cloud docs now brand the platform Gemini Enterprise Agent Platform. Vertex AI Platform appears in Google's SOC 1/2/3 and ISO 27001 scope; public GSU price list not verified at review.

Where they diverge

Deployment differentiation

Only Azure OpenAI Provisioned Throughput (PTU)

Global provisionedData-zone provisionedRegional provisioned

Both

No overlap in deployment models.

Only Vertex AI Provisioned Throughput

Provisioned throughputSingle-zone provisioned throughput

Read the full profiles