Select your cookie preferences

We use essential cookies and similar tools that are necessary to provide our site and services. We use performance cookies to collect anonymous statistics, so we can understand how customers use our site and make improvements. Essential cookies cannot be deactivated, but you can choose “Customize” or “Decline” to decline performance cookies.

If you agree, AWS and approved third parties will also use cookies to provide useful site features, remember your preferences, and display relevant content, including relevant advertising. To accept or decline all non-essential cookies, choose “Accept” or “Decline.” To make more detailed choices, choose “Customize.”

Use an inference profile in model invocation

Focus mode
Use an inference profile in model invocation - Amazon Bedrock

You can use a cross region inference profile in place of a foundation model to route requests to multiple Regions. To track costs and usage for a model, in one or multiple Regions, you can use an application inference profile. To learn how to use an inference profile when running model inference, choose the tab for your preferred method, and then follow the steps:

Console

In the console, the only inference profile you can use is the US Anthropic Claude 3 Opus inference profile in the US East (N. Virginia) region.

To use this inference profile, switch to the US East (N. Virginia) region. Do one of the following and select the Anthropic Claude 3 Opus model and Cross region inference as the Throughput when you reach the step to select a model:

API

You can use an inference profile when running inference from any Region that is included in it with the following API operations:

Note

If you're using a cross-region (system-defined) inference profile, you can use either the ARN or the ID of the inference profile.

In the console, the only inference profile you can use is the US Anthropic Claude 3 Opus inference profile in the US East (N. Virginia) region.

To use this inference profile, switch to the US East (N. Virginia) region. Do one of the following and select the Anthropic Claude 3 Opus model and Cross region inference as the Throughput when you reach the step to select a model:

PrivacySite termsCookie preferences
© 2025, Amazon Web Services, Inc. or its affiliates. All rights reserved.