← Back to all announcements
★★★☆☆ 30/06/2026

Claude Opus 4.8 is now available in AWS GovCloud (US)

Anthropic's most capable model now runs inside GovCloud, bringing 1M-token reasoning and agentic autonomy to regulated federal workloads.

View original announcement →

Visual Summary

graph TD A{{Claude Opus 4.8 in GovCloud}}:::announced B(Amazon Bedrock):::compute C([Extended Reasoning]):::feature D([Agentic Coding]):::feature E([Computer Use]):::feature F(Bedrock Guardrails):::compute G(Knowledge Bases):::storage H((Gov Agencies)):::external I([Prompt Caching]):::feature J([1M Token Context]):::feature H ==>|"invokes"| A A ==>|"served via"| B A -->|"enables"| C A -->|"performs"| D A -->|"supports"| E B -->|"enforces"| F B -->|"grounds with"| G A -.->|"optimizes"| I A -->|"processes"| J classDef announced fill:#ff9900,stroke:#ec7211,color:#fff,font-weight:bold classDef compute fill:#e3f2fd,stroke:#1565c0,color:#1565c0 classDef storage fill:#e8f5e9,stroke:#2e7d32,color:#2e7d32 classDef feature fill:#fff3e0,stroke:#e65100,color:#e65100 classDef external fill:#f5f5f5,stroke:#616161,color:#616161

What's New

Anthropic's Claude Opus 4.8, its most capable generally available model to date, is now accessible within AWS GovCloud (US) via Amazon Bedrock. The model delivers significant advances in agentic coding, long-running autonomous task execution, and professional knowledge work, making it suitable for production AI applications in regulated and government environments. This expansion brings enterprise-grade AI capabilities to workloads requiring strict data residency and compliance controls inherent to the GovCloud infrastructure.

How It Works

  • Claude Opus 4.8 is accessed through Amazon Bedrock using the model ID anthropic.claude-opus-4-8, supporting both the bedrock-runtime and bedrock-mantle endpoints for programmatic invocation via the Converse, InvokeModel, and Messages APIs.
  • The model features a 1 million token context window and supports up to 128K output tokens, enabling it to ingest and reason over entire codebases or large document corpora in a single session.
  • Extended reasoning is natively supported, allowing the model to perform multi-step deliberation before producing outputs, which is critical for complex agentic workflows.
  • Prompt caching is supported with a minimum of 4,096 tokens per cache checkpoint, up to 4 checkpoints per request, with TTLs of 5 minutes or 1 hour, reducing latency and cost for repetitive long-context calls.
  • Computer use capability is available via both bedrock-runtime and bedrock-mantle endpoints using the computer-use-2025-11-24 tool type, enabling GUI-level automation tasks.
  • AWS Bedrock Guardrails, Knowledge Bases, Agents, Flows, and Prompt Management are all supported, allowing teams to build governed, retrieval-augmented, and orchestrated AI pipelines around the model.
  • Enabling the model in GovCloud requires first accepting the EULA in a linked standard AWS account (us-east-1 or us-west-2) via the console or AWS CLI, then activating it through the GovCloud Model Access page.

Why It's Important

  • Government agencies, defense contractors, and regulated enterprises can now leverage Anthropic's most capable model without moving sensitive data outside the FedRAMP-authorized GovCloud boundary.
  • The model's ability to perform longer autonomous runs and recover from errors without human intervention reduces the need for constant oversight in complex, multi-step workflows—directly lowering operational overhead.
  • Deep codebase comprehension and long-session context retention make it viable for large-scale software modernization projects common in government IT, such as legacy system migration and security auditing.
  • Native support for Bedrock Guardrails and Knowledge Bases means organizations can enforce content policies and ground the model in authoritative internal data sources, addressing compliance and accuracy requirements simultaneously.
  • The 1M token context window allows synthesis across entire policy documents, legal corpora, or technical specifications in a single call, dramatically accelerating knowledge work that previously required manual chunking and aggregation.

How It's Different

  • Unlike earlier Claude models, Opus 4.8 is explicitly optimized for agentic resilience—it actively navigates around obstacles, self-corrects errors, and makes autonomous judgments about when to escalate versus proceed, rather than stalling or failing silently.
  • The 1M token context window is substantially larger than most competing models available in GovCloud, enabling true long-document synthesis without retrieval-augmented workarounds for many use cases.
  • Extended reasoning support differentiates it from lighter Claude variants (e.g., Haiku, Sonnet) by enabling deeper, chain-of-thought-style deliberation before output generation, improving accuracy on complex analytical tasks.
  • Computer use support—available in GovCloud—sets it apart from models limited to text and code, enabling end-to-end automation of GUI-based workflows that are common in legacy government systems.
  • Prompt caching with multi-checkpoint support reduces inference costs for long-context, repetitive workloads compared to models that require full context re-processing on every call.
  • Access through Amazon Bedrock's unified API means teams can switch between inference routing modes (In-Region, Geo, Global) without code changes, offering flexibility that direct API integrations with model providers cannot match.

When to Prefer It

  • Choose Claude Opus 4.8 in GovCloud when building autonomous AI agents that must complete multi-step tasks—such as software deployment pipelines, procurement workflows, or incident response automation—with minimal human intervention.
  • Prefer it for large-scale code modernization or security review tasks where the model needs to read, understand, and edit across an entire repository in a single coherent session.
  • Use it for knowledge-intensive deliverables—policy analysis, contract review, regulatory compliance summaries—where the model must synthesize across hundreds of pages and produce structured, reviewable outputs.
  • It is the right choice when your workload is subject to FedRAMP, ITAR, or other federal compliance frameworks that prohibit data from leaving GovCloud boundaries, and you need a top-tier model within those constraints.
  • Select it over lighter models (Haiku, Sonnet) when task complexity, reasoning depth, or output quality is the primary concern and cost-per-token is a secondary consideration.
  • Prefer it when your application requires computer use automation against legacy GUI-based systems that lack APIs, particularly in government IT modernization contexts.

Availability

  • Status: Generally Available (GA) in AWS GovCloud (US) as of June 30, 2026; model launch date was May 28, 2026 in commercial regions.
  • GovCloud Regions: Available in AWS GovCloud (US) regions; In-Region inference keeps data strictly within the GovCloud boundary.
  • Commercial Regions: Also available across US, EU, Japan, and Australia geographies via Geo inference profiles, and globally via the Global inference profile for non-GovCloud workloads.
  • Endpoints: Supported on both bedrock-runtime (model ID: anthropic.claude-opus-4-8) and bedrock-mantle endpoints.
  • Pricing: Refer to the Amazon Bedrock Pricing page; no specific per-token pricing was published in the announcement, but prompt caching is supported to reduce costs on long-context workloads.
  • GovCloud Enablement Requirement: EULA must be accepted in a linked standard AWS account (us-east-1 or us-west-2) before the model can be activated in the GovCloud account; entitlement propagation may take a few minutes.
  • Model Lifecycle: Active, with no announced End-of-Life (EOL) date.
  • Context / Output Limits: 1M token context window; 128K max output tokens; knowledge cutoff is January 2026.

Tags

Servicesother-aws
Typeregion-expansionga-launch
Conceptsgenaillmagentic-aicoding-assistant
Use Casesenterprisedeveloper-toolsgovernment
Providersanthropic
GeographyAMERICAS

Related Resources

AI Radar AWS

AWS AI/ML news — curated, researched, explained

An automated intelligence platform that curates, researches, and analyzes AWS AI/ML/GenAI announcements daily. Every report is backed by real research — the system reads linked blog posts and documentation to provide accurate, in-depth analysis.

How Each Report Is Generated

  1. Collection — Daily monitoring of the AWS "What's New" RSS feed
  2. Filtering — AI-powered relevance detection for AI/ML/GenAI topics
  3. Taxonomy Tagging — LLM-based classification across 6 dimensions
  4. Importance Scoring — Point-based system with tag bonuses (1-5 stars)
  5. Research Phase — Follows links to blog posts and documentation
  6. Report Generation — Claude Sonnet produces structured 6-section analysis
  7. Visual Summary — Claude Opus generates Mermaid diagrams for key items
  8. Publishing — Static website rebuilt and deployed via CloudFront

Features

  • Faceted filtering by service, type, concept, and more
  • Multi-dimensional taxonomy with 80+ tags across 6 dimensions
  • Geographic availability badges (Global, APJ, EMEA, AMER) with filtering
  • Timeline visualization of announcement volume
  • PDF export for offline reading
  • Mermaid visual summaries for key announcements
  • Daily automated updates — no manual curation
What makes this different: Each report involves a dedicated research phase where the system reads linked blog posts and AWS documentation pages. This produces analysis that goes beyond the original announcement text.

Technology

Built with Python, AWS Lambda, Amazon Bedrock (Claude Sonnet 4.6, Opus 4.6, Haiku 4.5), S3, CloudFront, WAF, EventBridge, and CDK.

Open Source

This project is open source. Fork it, customize it for your needs, and deploy your own instance.
📦 github.com/bbonik/ai-radar-aws

How Importance Scoring Works

Each announcement receives a point score based on multiple factors. The total score maps to a 1-5 star rating:

1★ < 2 pts 2★ ≥ 2 pts 3★ ≥ 3.5 pts 4★ ≥ 5 pts 5★ ≥ 6.5 pts

Point Breakdown

FactorPointsWhen
Core AI service (Bedrock, AgentCore, SageMaker AI)+4Service named in title
Key AI service (SageMaker, Kiro, QuickSight)+2Service named in title
Other AI-related service+1Default
Blog post link+3Link to aws.amazon.com/blogs/
GitHub samples link+2Link to github.com/aws*
Documentation link+1Link to docs.aws.amazon.com/
New model+1.5Tagged as "new-model"
New service+1Tagged as "new-service"
New feature+0.5Tagged as "new-feature"
Anthropic / OpenAI provider+2Provider explicitly mentioned
Instance / notebook announcement-2Hardware/capacity, not feature
Performance / pricing / security-0.5Incremental updates
Region expansion to APJ+1Expands to Asia Pacific
Region expansion (non-APJ only)-1.5Only expands to other regions

Geographic Relevance Badges

Each announcement card shows a small badge indicating whether the feature is available in your region:

🌐 Global Available in all regions
🌏 APJ Asia Pacific
🌍 EMEA Europe / Middle East / Africa
🌎 AMER Americas (US, Canada, South America)
No badge Geography unknown
How geography is detected: The system detects ALL geographies mentioned in each announcement. If the text mentions specific regions (Tokyo, Frankfurt, Oregon, etc.), the corresponding geography badges are shown. If it says "all regions" or is a new feature with no region specified, it gets the Global badge. Geography is also filterable — click a geo chip to see only announcements available in that region.