Vendors · Comparison
Amazon Bedrock Provisioned Throughput vs Vertex AI Provisioned Throughput
Side-by-side comparison of Amazon Bedrock Provisioned Throughput and Vertex AI Provisioned Throughput for single-tenant LLM hosting. Deployment options, compliance, pricing, and operational fit compared.
Amazon Bedrock Provisioned Throughput
Provisioned Throughput reserves per-model capacity in Model Units (MUs), billed hourly with no-commitment, 1-month, or 6-month terms; required for serving customised models. AWS states Bedrock content is stored at rest in the Region of use and not shared with model providers. Example per-MU rates are published for some models; most quotes require an AWS account team.
Vertex AI Provisioned Throughput
Provisioned Throughput reserves Gemini and partner-model capacity in generative AI scale units (GSUs) on fixed-cost terms, including 1-week options for select models. Google Cloud docs now brand the platform Gemini Enterprise Agent Platform. Vertex AI Platform appears in Google's SOC 1/2/3 and ISO 27001 scope; public GSU price list not verified at review.
From the data
Key differences
- Category Both are Hyperscaler dedicated offerings.
- Compliance Both document SOC 2 Type II, SOC 3, ISO 27001 and GDPR DPA. Only Amazon Bedrock Provisioned Throughput documents HIPAA BAA, C5 and FedRAMP.
- Data residency Amazon Bedrock Provisioned Throughput has a verified residency guarantee; Vertex AI Provisioned Throughput's residency guarantee is not verified.
- Jurisdiction Both parent companies are under US jurisdiction.
- Pricing Amazon Bedrock Provisioned Throughput: Per Model Unit (MU) per hour; no-commitment, 1-month, or 6-month terms (no published pricing). Vertex AI Provisioned Throughput: Per GSU (generative AI scale unit) fixed-cost term; 1-week terms for select models (publication not verified).
Generated from the verified vendor data below; sources and dates on each vendor profile.
Side by side
Capabilities compared
| Amazon Bedrock Provisioned Throughput | Vertex AI Provisioned Throughput | |
|---|---|---|
| Category | Hyperscaler dedicated | Hyperscaler dedicated |
| Deployment models | Provisioned throughputCustom model (provisioned) | Provisioned throughputSingle-zone provisioned throughput |
| Regions | US, Europe, Asia-Pacific | US, Europe, Asia-Pacific |
| Compliance | SOC 2 Type IISOC 3ISO 27001HIPAA BAAC5FedRAMPGDPR DPA | SOC 2 Type IISOC 3ISO 27001GDPR DPA |
| Pricing model | Per Model Unit (MU) per hour; no-commitment, 1-month, or 6-month terms | Per GSU (generative AI scale unit) fixed-cost term; 1-week terms for select models |
| Public pricing | No | Not verified |
| Residency guarantee | Yes | Not verified |
| Parent jurisdiction | US | US |
| Analyst note | Provisioned Throughput reserves per-model capacity in Model Units (MUs), billed hourly with no-commitment, 1-month, or 6-month terms; required for serving customised models. AWS states Bedrock content is stored at rest in the Region of use and not shared with model providers. Example per-MU rates are published for some models; most quotes require an AWS account team. | Provisioned Throughput reserves Gemini and partner-model capacity in generative AI scale units (GSUs) on fixed-cost terms, including 1-week options for select models. Google Cloud docs now brand the platform Gemini Enterprise Agent Platform. Vertex AI Platform appears in Google's SOC 1/2/3 and ISO 27001 scope; public GSU price list not verified at review. |
Where they diverge
Deployment differentiation
Only Amazon Bedrock Provisioned Throughput
Both
Only Vertex AI Provisioned Throughput
Read the full profiles