Introducing Anthropic models on Amazon Bedrock for in-region inference in Seoul and Singapore Amazon Bedrock now supports in-region inference for Anthropic's Claude Opus 5 and Claude Sonnet 5 in the Asia Pacific (Seoul) Region (ap-northeast-2) and Claude Sonnet 5 in the Asia Pacific (Singapore) Region (ap-southeast-1) on the bedrock-runtime endpoint, keeping inference requests and data within the Region called. The models are invoked with direct model IDs such as anthropic.claude-opus-5 and anthropic.claude-sonnet-5, and support Anthropic's Messages API plus the Amazon Bedrock InvokeModel and Converse APIs, Amazon Bedrock Guardrails, and intelligent prompt routing. In-region inference has no routing layer, so throughput is bounded by each Region's capacity and subject to per-Region service quotas, with billing at standard on-demand pricing for the Region called. Artificial Intelligence https://aws.amazon.com/blogs/machine-learning/ Introducing Anthropic models on Amazon Bedrock for in-region inference in Seoul and Singapore Amazon Bedrock now supports the Anthropic Claude models: Claude Opus 5 https://aws.amazon.com/blogs/machine-learning/introducing-claude-opus-5-on-aws-anthropics-most-capable-opus-model/ and Claude Sonnet 5 https://aws.amazon.com/blogs/machine-learning/introducing-claude-sonnet-5-on-aws-anthropics-most-capable-sonnet-model/ in Seoul and Claude Sonnet 5 https://aws.amazon.com/blogs/machine-learning/introducing-claude-sonnet-5-on-aws-anthropics-most-capable-sonnet-model/ in Singapore with in-region inference on the bedrock-runtime endpoint. If you have local data processing requirements in South Korea or Singapore, for example, in financial services, healthcare, and the public sector, you can now use these Anthropic models at scale. Amazon Bedrock processes inference requests and data within the Region you call. The processing does not leave the Region. In this post, we walk through how in-region inference works from the Asia Pacific Seoul Region ap-northeast-2 and Asia Pacific Singapore Region ap-southeast-1 using the bedrock-runtime endpoint. We also show how to get started from the Amazon Bedrock console and with code, using the Amazon Bedrock Converse API, the InvokeModel API, and the Anthropic Messages API. In-region inference To help you meet strict data residency requirements for your AI applications, Amazon Bedrock offers in-region inference. Your request is processed entirely within the single AWS Region you specify, and it does not leave that Region. Use this when you have a need for strict single-Region data processing. Unlike cross-Region inference profiles, there is no routing layer. The request you send to the Seoul ap-northeast-2 or Singapore ap-southeast-1 Region is served by that Region alone. Your input prompts and output results stay within it for the full lifecycle of the request. In exchange, your throughput is bounded by that Region’s capacity. Requests are subject to per-Region service quotas. Billing follows standard on-demand pricing for the Region you call. Quota consumption, Amazon CloudWatch metrics, and AWS CloudTrail log entries are all scoped to that same Region. There is no source-versus-destination distinction to account for in your monitoring. In-region inference for Claude Sonnet 5 and Claude Opus 5 in Seoul and Claude Sonnet 5 in Singapore is available on the bedrock-runtime endpoint. For new applications, we recommend the bedrock-runtime endpoint. You invoke it with the direct model ID, for example, anthropic.claude-opus-5 or anthropic.claude-sonnet-5 . It supports Anthropic’s Messages https://docs.aws.amazon.com/bedrock/latest/userguide/inference-messages-api.html API and the Amazon Bedrock InvokeModel https://docs.aws.amazon.com/bedrock/latest/userguide/inference-api.html and Converse https://docs.aws.amazon.com/bedrock/latest/userguide/conversation-inference.html APIs, along with Amazon Bedrock features such as Amazon Bedrock Guardrails https://docs.aws.amazon.com/bedrock/latest/userguide/guardrails.html and intelligent prompt routing https://docs.aws.amazon.com/bedrock/latest/userguide/prompt-routing.html . Access Claude models from the Amazon Bedrock console You can access Claude models in the text playground in the Amazon Bedrock console, which requires no coding or SDK setup. You can send prompts, adjust inference parameters, and switch between variants to get a feel for each model before you integrate the API. 1. Open the Amazon Bedrock console https://console.aws.amazon.com/bedrock/ in a Region that you want to use as a source. 2. In the navigation pane, under Test , choose Playground . 3. Choose Select model in the middle of the page. 4. Search for anthropic.claude-opus-5 , select On-Demand under Inference , and choose Apply . 5. Enter a prompt and choose Run to generate a response. Call Claude models with the Anthropic Messages API and Amazon Bedrock InvokeModel and Converse API You can access Anthropic’s Claude Opus 5 or Claude Sonnet 5 programmatically with Seoul in-region inference using the Anthropic Messages API https://docs.aws.amazon.com/bedrock/latest/userguide/model-parameters-anthropic-claude-messages.html?trk=d8ec3b19-0f37-4f8c-8c12-189f913e205c&sc channel=el on the bedrock-runtime through Anthropic SDK or keep using the Invoke https://docs.aws.amazon.com/bedrock/latest/userguide/inference-api.html and Converse API https://docs.aws.amazon.com/bedrock/latest/userguide/conversation-inference.html?trk=d8ec3b19-0f37-4f8c-8c12-189f913e205c&sc channel=el on bedrock-runtime through the AWS Command Line Interface AWS CLI https://aws.amazon.com/cli/?trk=769a1a2b-8c19-4976-9c45-b6b1226c7d20&sc channel=el and AWS SDK https://aws.amazon.com/developer/tools/?trk=769a1a2b-8c19-4976-9c45-b6b1226c7d20&sc channel=el . Prerequisites 1. Active AWS account with Amazon Bedrock access. 2. AWS CLI installed and configured. 3. Python 3.8+. 4. Boto3 installed: pip install boto3 . 5. Anthropic SDK installed: pip install anthropic . 6. The Amazon Bedrock Token Generator for Amazon Bedrock authentication installed: pip install aws bedrock token generator . Here’s a quick example using the AWS SDK for Python Boto3 with the InvokeModel API: You can also use the Amazon Bedrock Converse API for a unified multi-model experience: You can also use the Anthropic Messages API through the anthropic SDK package for a streamlined experience: You can monitor usage, performance, and costs through CloudWatch https://docs.aws.amazon.com/bedrock/latest/userguide/monitoring.html and AWS Cost Explorer https://aws.amazon.com/aws-cost-management/aws-cost-explorer/ to scale your applications as demand grows. Conclusion With the launch of Anthropic’s Claude Opus 5 and Claude Sonnet 5 on Amazon Bedrock with in-region inference in Seoul, and Claude Sonnet 5 with in-region inference in Singapore, you can now build generative AI applications that have strict data residency requirements. This keeps inference within the Region you want. We are excited about this launch and look forward to seeing how you use these capabilities to accelerate innovation and deliver impactful AI-powered experiences across the Region. For the most current information about model availability in each Region, see Regional availability by models https://docs.aws.amazon.com/bedrock/latest/userguide/models-region-compatibility.html model-regions-anthropic in the Amazon Bedrock User Guide.