← Back to all announcements
★★★★☆ 30/07/2026

Grok 4.3 from xAI is now available on Amazon Bedrock in AWS GovCloud (US-West)

Government and regulated-industry teams can now run xAI's reasoning-optimized Grok 4.3 inside AWS GovCloud with configurable cost controls.

View original announcement →

Visual Summary

graph TD A{{Grok 4.3 on GovCloud}}:::announced B(Amazon Bedrock):::compute C([Mantle Inference Engine]):::feature D([Configurable Reasoning]):::feature E([Tool Use & Structured Output]):::feature F((Gov/Enterprise Users)):::external G(AWS GovCloud US-West):::compute H([In-Region Routing]):::feature I([Agentic Workflows]):::feature F ==>|"invokes"| A A ==>|"runs on"| B B -->|"deployed in"| G A -->|"powered by"| C A -->|"supports"| D A -->|"enables"| E G -.->|"enforces"| H E -->|"builds"| I classDef announced fill:#ff9900,stroke:#ec7211,color:#fff,font-weight:bold classDef compute fill:#e3f2fd,stroke:#1565c0,color:#1565c0 classDef storage fill:#e8f5e9,stroke:#2e7d32,color:#2e7d32 classDef feature fill:#fff3e0,stroke:#e65100,color:#e65100 classDef external fill:#f5f5f5,stroke:#616161,color:#616161

What's New

xAI's Grok 4.3 reasoning model is now available on Amazon Bedrock in AWS GovCloud (US-West), marking xAI's first entry as a model provider in the GovCloud environment. The model is a reasoning-first large language model with configurable reasoning effort levels and strong agentic capabilities, running on Mantle, a new inference engine in Amazon Bedrock optimized for price-performance. This launch expands the generative AI options available to government and regulated-industry customers building compliant, high-stakes AI applications.

How It Works

  • Configurable Reasoning Effort: Grok 4.3 supports four reasoning effort levels—none, low, medium, and high—allowing developers to trade off latency and cost against reasoning depth depending on the task complexity.
  • Mantle Inference Engine: The model runs on Mantle, a new Amazon Bedrock inference engine purpose-built for price-performance, providing efficient token throughput for high-volume workloads.
  • Tool Use and Structured Output: Grok 4.3 natively supports tool calling, structured output (e.g., JSON schemas), and response streaming, enabling reliable integration into agentic pipelines and API-driven workflows.
  • Multi-Turn and Instruction Following: The model is designed for strong instruction adherence across multi-turn conversations, making it suitable for complex, stateful dialogue and workflow automation.
  • GovCloud Access Process: Enabling Grok 4.3 in GovCloud requires first accepting the EULA in a linked standard AWS account (us-east-1 or us-west-2) via the Console or AWS CLI, then enabling the model through the GovCloud Model Access page.
  • Inference Routing Options: GovCloud users can use In-Region routing to ensure data never leaves the AWS GovCloud (US-West) region, satisfying strict data residency and compliance requirements.

Why It's Important

  • Expands GovCloud AI Model Diversity: xAI becomes a new model provider on Amazon Bedrock in GovCloud, giving government agencies and regulated enterprises access to a broader portfolio of frontier models within a compliant boundary.
  • Addresses Regulated Workload Needs: Government customers handling sensitive data—such as legal case research, financial document analysis, or citizen support—can now leverage a capable reasoning model without sacrificing data sovereignty.
  • Cost-Effective High-Volume Inference: Token efficiency combined with the Mantle engine makes Grok 4.3 viable for large-scale deployments where inference cost is a primary constraint, such as document processing pipelines.
  • Agentic AI in Secure Environments: Native tool use and structured output support enables the construction of autonomous agents within GovCloud, unlocking automation for workflows that previously required manual intervention.
  • Reasoning Configurability Reduces Waste: The ability to dial down reasoning effort for simpler tasks means organizations are not paying for heavy compute when lightweight responses suffice, improving operational efficiency.

How It's Different

  • Reasoning-First Architecture with Dial Control: Unlike many models that apply fixed reasoning depth, Grok 4.3 exposes explicit reasoning effort controls (none/low/medium/high), giving developers fine-grained cost and latency management not commonly available in other Bedrock models.
  • Runs on Mantle, Not Standard Bedrock Runtime: Grok 4.3 is among the first models to run on Mantle, Amazon Bedrock's new inference engine optimized specifically for price-performance, distinguishing it from models on the standard runtime.
  • First xAI Model in GovCloud: No other xAI model has previously been available in AWS GovCloud, making this a unique option for customers who require xAI's capabilities within a FedRAMP-authorized environment.
  • Token Efficiency Focus: Grok 4.3 is specifically designed for token efficiency at scale, which differentiates it from larger, more expensive frontier models that may offer higher raw capability but at significantly greater cost per token.
  • Broad Enterprise Workflow Coverage: The model is explicitly tuned for enterprise verticals (legal, financial, customer support, web development) rather than being a general-purpose model, offering more consistent out-of-the-box performance in those domains.

When to Prefer It

  • Government and Public Sector AI Applications: Choose Grok 4.3 on GovCloud when building AI solutions for federal, state, or local government agencies that require data to remain within FedRAMP-authorized infrastructure.
  • Legal and Compliance Research: Use Grok 4.3 when the workload involves case law research, contract review, or regulatory document analysis, where its reasoning capabilities and instruction following provide reliable, structured outputs.
  • Financial Document Q&A: Prefer Grok 4.3 for applications that extract insights from financial reports, filings, or audit documents, especially when accuracy and structured output are critical.
  • High-Volume Inference at Controlled Cost: Select Grok 4.3 when deploying at scale (e.g., batch document processing, high-traffic customer support bots) where token efficiency and Mantle's price-performance characteristics reduce total inference spend.
  • Agentic Workflows Requiring Tool Use: Use Grok 4.3 when building multi-step agents that call external APIs, query databases, or execute structured actions, leveraging its native tool calling and structured output support.
  • Variable Complexity Tasks: Prefer Grok 4.3 when your application handles a mix of simple and complex queries, allowing you to dynamically adjust reasoning effort per request to optimize cost without degrading quality on harder tasks.
  • Multi-Turn Conversational Applications: Choose Grok 4.3 for chat, search, and dialogue systems that require consistent context retention and instruction adherence across long conversation histories.

Availability

  • Status: Generally Available (GA) as of July 30, 2026.
  • Primary Region: AWS GovCloud (US-West); this announcement specifically highlights GovCloud availability as the new addition.
  • Additional Regions: Grok 4.3 is also available in other Amazon Bedrock commercial regions; refer to the official regional availability page for the full list.
  • Access Requirement: GovCloud customers must first accept the model EULA in a linked standard AWS account (us-east-1 or us-west-2) before enabling the model in their GovCloud account via the Model Access page.
  • Inference Engine: Runs on Mantle, Amazon Bedrock's new price-performance inference engine, with support for tool calling, structured output, and response streaming.
  • Pricing: Specific pricing is not disclosed in the announcement; consult the Amazon Bedrock pricing page for current on-demand token rates for Grok 4.3.
  • Limitations: GovCloud access requires a linked standard AWS account for EULA acceptance; entitlement propagation may take a few minutes after enablement.

Tags

Servicesbedrock
Typenew-modelregion-expansionga-launch
Conceptsgenaillmagentic-aiinferenceconversational-ai
Use Casesenterprisegovernment
GeographyAMERICAS

Related Resources

AI Radar AWS

AWS AI/ML news — curated, researched, explained

An automated intelligence platform that curates, researches, and analyzes AWS AI/ML/GenAI announcements daily. Every report is backed by real research — the system reads linked blog posts and documentation to provide accurate, in-depth analysis.

How Each Report Is Generated

  1. Collection — Daily monitoring of the AWS "What's New" RSS feed
  2. Filtering — AI-powered relevance detection for AI/ML/GenAI topics
  3. Taxonomy Tagging — LLM-based classification across 6 dimensions
  4. Importance Scoring — Point-based system with tag bonuses (1-5 stars)
  5. Research Phase — Follows links to blog posts and documentation
  6. Report Generation — Claude Sonnet produces structured 6-section analysis
  7. Visual Summary — Claude Opus generates Mermaid diagrams for key items
  8. Publishing — Static website rebuilt and deployed via CloudFront

Features

  • Faceted filtering by service, type, concept, and more
  • Multi-dimensional taxonomy with 80+ tags across 6 dimensions
  • Geographic availability badges (Global, APJ, EMEA, AMER) with filtering
  • Timeline visualization of announcement volume
  • PDF export for offline reading
  • Mermaid visual summaries for key announcements
  • Daily automated updates — no manual curation
What makes this different: Each report involves a dedicated research phase where the system reads linked blog posts and AWS documentation pages. This produces analysis that goes beyond the original announcement text.

Technology

Built with Python, AWS Lambda, Amazon Bedrock (Claude Sonnet 4.6, Opus 4.6, Haiku 4.5), S3, CloudFront, WAF, EventBridge, and CDK.

Open Source

This project is open source. Fork it, customize it for your needs, and deploy your own instance.
📦 github.com/bbonik/ai-radar-aws

How Importance Scoring Works

Each announcement receives a point score based on multiple factors. The total score maps to a 1-5 star rating:

1★ < 2 pts 2★ ≥ 2 pts 3★ ≥ 3.5 pts 4★ ≥ 5 pts 5★ ≥ 6.5 pts

Point Breakdown

FactorPointsWhen
Core AI service (Bedrock, AgentCore, SageMaker AI)+4Service named in title
Key AI service (SageMaker, Kiro, QuickSight)+2Service named in title
Other AI-related service+1Default
Blog post link+3Link to aws.amazon.com/blogs/
GitHub samples link+2Link to github.com/aws*
Documentation link+1Link to docs.aws.amazon.com/
New model+1.5Tagged as "new-model"
New service+1Tagged as "new-service"
New feature+0.5Tagged as "new-feature"
Anthropic / OpenAI provider+2Provider explicitly mentioned
Instance / notebook announcement-2Hardware/capacity, not feature
Performance / pricing / security-0.5Incremental updates
Region expansion to APJ+1Expands to Asia Pacific
Region expansion (non-APJ only)-1.5Only expands to other regions

Geographic Relevance Badges

Each announcement card shows a small badge indicating whether the feature is available in your region:

🌐 Global Available in all regions
🌏 APJ Asia Pacific
🌍 EMEA Europe / Middle East / Africa
🌎 AMER Americas (US, Canada, South America)
No badge Geography unknown
How geography is detected: The system detects ALL geographies mentioned in each announcement. If the text mentions specific regions (Tokyo, Frankfurt, Oregon, etc.), the corresponding geography badges are shown. If it says "all regions" or is a new feature with no region specified, it gets the Global badge. Geography is also filterable — click a geo chip to see only announcements available in that region.