singletenant.ai

Vendors · Comparison

Amazon Bedrock Provisioned Throughput vs Together AI

Side-by-side comparison of Amazon Bedrock Provisioned Throughput and Together AI for single-tenant LLM hosting. Deployment options, compliance, pricing, and operational fit compared.

Amazon Bedrock Provisioned Throughput

Provisioned Throughput reserves per-model capacity in Model Units (MUs), billed hourly with no-commitment, 1-month, or 6-month terms; required for serving customised models. AWS states Bedrock content is stored at rest in the Region of use and not shared with model providers. Example per-MU rates are published for some models; most quotes require an AWS account team.

Together AI

Dedicated endpoints are explicitly single-tenant GPU instances with public per-hour pricing. Terms of service make no geographic storage commitment, so residency guarantee is recorded as none.

From the data

Key differences

  • Category Amazon Bedrock Provisioned Throughput is a Hyperscaler dedicated offering; Together AI is Dedicated GPU.
  • Compliance Both document SOC 2 Type II and HIPAA BAA. Only Amazon Bedrock Provisioned Throughput documents SOC 3, ISO 27001, C5, FedRAMP and GDPR DPA.
  • Data residency Amazon Bedrock Provisioned Throughput has a verified residency guarantee; Together AI offers no residency guarantee (verified).
  • Jurisdiction Both parent companies are under US jurisdiction.
  • Pricing Amazon Bedrock Provisioned Throughput: Per Model Unit (MU) per hour; no-commitment, 1-month, or 6-month terms (no published pricing). Together AI: Per 1M tokens serverless; per GPU-hour for dedicated endpoints and clusters (published pricing).

Generated from the verified vendor data below; sources and dates on each vendor profile.

Side by side

Capabilities compared

Amazon Bedrock Provisioned Throughput Together AI
Category Hyperscaler dedicated Dedicated GPU
Deployment models Provisioned throughputCustom model (provisioned) ServerlessDedicated endpointGPU clusters
Regions US, Europe, Asia-Pacific Not verified
Compliance SOC 2 Type IISOC 3ISO 27001HIPAA BAAC5FedRAMPGDPR DPA SOC 2 Type IIHIPAA BAA
Pricing model Per Model Unit (MU) per hour; no-commitment, 1-month, or 6-month terms Per 1M tokens serverless; per GPU-hour for dedicated endpoints and clusters
Public pricing No Yes
Residency guarantee Yes No
Parent jurisdiction US US
Analyst note Provisioned Throughput reserves per-model capacity in Model Units (MUs), billed hourly with no-commitment, 1-month, or 6-month terms; required for serving customised models. AWS states Bedrock content is stored at rest in the Region of use and not shared with model providers. Example per-MU rates are published for some models; most quotes require an AWS account team. Dedicated endpoints are explicitly single-tenant GPU instances with public per-hour pricing. Terms of service make no geographic storage commitment, so residency guarantee is recorded as none.

Where they diverge

Deployment differentiation

Only Amazon Bedrock Provisioned Throughput

Provisioned throughputCustom model (provisioned)

Both

No overlap in deployment models.

Only Together AI

ServerlessDedicated endpointGPU clusters

Read the full profiles