singletenant.ai

Vendors · Comparison

Amazon Bedrock Provisioned Throughput vs Vertex AI Provisioned Throughput

Side-by-side comparison of Amazon Bedrock Provisioned Throughput and Vertex AI Provisioned Throughput for single-tenant LLM hosting. Deployment options, compliance, pricing, and operational fit compared.

Amazon Bedrock Provisioned Throughput

Provisioned Throughput reserves per-model capacity in Model Units (MUs), billed hourly with no-commitment, 1-month, or 6-month terms; required for serving customised models. AWS states Bedrock content is stored at rest in the Region of use and not shared with model providers. Example per-MU rates are published for some models; most quotes require an AWS account team.

Vertex AI Provisioned Throughput

Provisioned Throughput reserves Gemini and partner-model capacity in generative AI scale units (GSUs) on fixed-cost terms, including 1-week options for select models. Google Cloud docs now brand the platform Gemini Enterprise Agent Platform. Vertex AI Platform appears in Google's SOC 1/2/3 and ISO 27001 scope; public GSU price list not verified at review.

From the data

Key differences

  • Category Both are Hyperscaler dedicated offerings.
  • Compliance Both document SOC 2 Type II, SOC 3, ISO 27001 and GDPR DPA. Only Amazon Bedrock Provisioned Throughput documents HIPAA BAA, C5 and FedRAMP.
  • Data residency Amazon Bedrock Provisioned Throughput has a verified residency guarantee; Vertex AI Provisioned Throughput's residency guarantee is not verified.
  • Jurisdiction Both parent companies are under US jurisdiction.
  • Pricing Amazon Bedrock Provisioned Throughput: Per Model Unit (MU) per hour; no-commitment, 1-month, or 6-month terms (no published pricing). Vertex AI Provisioned Throughput: Per GSU (generative AI scale unit) fixed-cost term; 1-week terms for select models (publication not verified).

Generated from the verified vendor data below; sources and dates on each vendor profile.

Side by side

Capabilities compared

Amazon Bedrock Provisioned Throughput Vertex AI Provisioned Throughput
Category Hyperscaler dedicated Hyperscaler dedicated
Deployment models Provisioned throughputCustom model (provisioned) Provisioned throughputSingle-zone provisioned throughput
Regions US, Europe, Asia-Pacific US, Europe, Asia-Pacific
Compliance SOC 2 Type IISOC 3ISO 27001HIPAA BAAC5FedRAMPGDPR DPA SOC 2 Type IISOC 3ISO 27001GDPR DPA
Pricing model Per Model Unit (MU) per hour; no-commitment, 1-month, or 6-month terms Per GSU (generative AI scale unit) fixed-cost term; 1-week terms for select models
Public pricing No Not verified
Residency guarantee Yes Not verified
Parent jurisdiction US US
Analyst note Provisioned Throughput reserves per-model capacity in Model Units (MUs), billed hourly with no-commitment, 1-month, or 6-month terms; required for serving customised models. AWS states Bedrock content is stored at rest in the Region of use and not shared with model providers. Example per-MU rates are published for some models; most quotes require an AWS account team. Provisioned Throughput reserves Gemini and partner-model capacity in generative AI scale units (GSUs) on fixed-cost terms, including 1-week options for select models. Google Cloud docs now brand the platform Gemini Enterprise Agent Platform. Vertex AI Platform appears in Google's SOC 1/2/3 and ISO 27001 scope; public GSU price list not verified at review.

Where they diverge

Deployment differentiation

Only Amazon Bedrock Provisioned Throughput

Custom model (provisioned)

Both

Provisioned throughput

Only Vertex AI Provisioned Throughput

Single-zone provisioned throughput

Read the full profiles