← Back to all announcements
★★★★★ 11/06/2026

OpenAI GPT-5.4 and GPT-5.5 models now available in US East (N. Virginia) on Amazon Bedrock

OpenAI's most powerful models land in AWS's largest region, bringing 272K-context agentic AI inside your existing AWS security perimeter.

View original announcement →

Visual Summary

graph TD A{{GPT-5.4 & GPT-5.5 on Bedrock}}:::announced B(Amazon Bedrock):::compute C([Responses API]):::feature D([272K Context Window]):::feature E([Tool Calling]):::feature F([Agentic Workflows]):::feature G((Enterprise Apps)):::external H(AWS IAM & VPC):::compute I([Multimodal Input]):::feature G ==>|"invokes"| B B ==>|"serves"| A A -->|"accessed via"| C A -->|"supports"| D A -->|"enables"| E A -->|"powers"| F A -->|"accepts"| I H -.->|"secures"| B classDef announced fill:#ff9900,stroke:#ec7211,color:#fff,font-weight:bold classDef compute fill:#e3f2fd,stroke:#1565c0,color:#1565c0 classDef storage fill:#e8f5e9,stroke:#2e7d32,color:#2e7d32 classDef feature fill:#fff3e0,stroke:#e65100,color:#e65100 classDef external fill:#f5f5f5,stroke:#616161,color:#616161

What's New

AWS has expanded the availability of OpenAI's GPT-5.4 and GPT-5.5 models to the US East (N. Virginia) Region on Amazon Bedrock. GPT-5.5 is OpenAI's most capable model, targeting advanced coding, research, agentic workflows, and software operation, while GPT-5.4 focuses on frontier reasoning, computer use, and long-context production workloads. Both models are accessible through the Responses API and join a growing portfolio of OpenAI models already available on Bedrock.

How It Works

  • Both GPT-5.4 and GPT-5.5 are served through Amazon Bedrock's managed inference infrastructure, allowing access via the Responses API without requiring direct OpenAI account credentials or separate API key management.
  • Both models support a 272K-token context window, enabling processing of very long documents, multi-turn conversations, and complex multi-step workflows within a single inference call.
  • Input modalities include both text and images, making the models suitable for multimodal tasks such as document analysis, screenshot interpretation, and visual reasoning.
  • The Responses API supports server-side and client-side tool calling, allowing models to invoke external tools, APIs, or code interpreters as part of agentic task execution.
  • Response streaming is supported, enabling real-time token delivery to end-user applications for lower perceived latency in interactive use cases.
  • GPT-5.5 is designed to understand open-ended goals, navigate ambiguity, and complete long-running tasks with minimal orchestration overhead from the developer.
  • GPT-5.4 is optimized for production workflows requiring reliable multi-step reasoning, software environment interaction, and output verification across complex business systems.
  • Additional OpenAI models available on Bedrock include GPT OSS Safeguard 120B, GPT OSS Safeguard 20B, GPT OSS 120B, and GPT OSS 20B, providing a full-stack option for both inference and safety guardrails.

Why It's Important

  • US East (N. Virginia) is AWS's largest and most interconnected region, meaning lower latency for the majority of North American enterprise workloads and broader VPC/networking integration options.
  • Enterprises can now access frontier OpenAI models within their existing AWS security perimeter, leveraging IAM, VPC endpoints, AWS PrivateLink, and CloudTrail audit logging without data leaving the AWS ecosystem.
  • The 272K-token context window directly addresses enterprise pain points around long-document processing, legal contract review, large codebase analysis, and extended agentic task chains.
  • Availability on Bedrock enables unified billing, consolidated cost management, and access to Bedrock features like model evaluation, guardrails, and prompt management alongside OpenAI models.
  • The inclusion of OpenAI OSS Safeguard models on the same platform allows teams to pair powerful frontier models with dedicated safety layers, simplifying responsible AI deployment.
  • Reduced orchestration requirements in GPT-5.5 lower the engineering burden for building autonomous agents, accelerating time-to-production for complex AI-driven workflows.

How It's Different

  • Unlike accessing OpenAI models directly via the OpenAI API, Bedrock hosting means no separate OpenAI account is required, and all traffic, logging, and access control flows through AWS-native tooling.
  • The 272K-token context window is substantially larger than many competing models available on Bedrock, supporting use cases that would otherwise require chunking or retrieval augmentation.
  • GPT-5.5's explicit design for minimal orchestration differentiates it from models that require heavy prompt engineering or external agent frameworks to complete multi-step tasks reliably.
  • GPT-5.4's computer use capability—operating software environments and verifying outputs—goes beyond standard text/code generation, enabling robotic process automation-style workflows driven by LLM reasoning.
  • The availability of companion OSS Safeguard models (20B and 120B) on the same platform is a differentiator versus standalone model providers, offering an integrated moderation layer at varying compute cost points.
  • Being available in additional AWS regions (not just a single endpoint) provides geographic redundancy and data residency flexibility that direct OpenAI API access does not natively offer.

When to Prefer It

  • Prefer GPT-5.5 when building autonomous agents that must handle ambiguous, open-ended goals across long workflows with minimal human-in-the-loop intervention or custom orchestration code.
  • Prefer GPT-5.4 for production business applications requiring reliable multi-step reasoning, tool use, and output verification—such as financial analysis pipelines, compliance checks, or ERP integrations.
  • Use either model when your organization requires all AI inference to remain within the AWS security boundary for compliance, data residency, or audit trail reasons.
  • Choose these models for long-document workflows—such as legal discovery, medical record summarization, or large codebase review—where the 272K-token context window eliminates the need for complex chunking strategies.
  • Prefer these models when you need multimodal input (text + images) for tasks like UI testing automation, document digitization, or visual data extraction within an agentic pipeline.
  • Use the GPT OSS Safeguard models in conjunction with GPT-5.4/5.5 when deploying customer-facing applications that require content moderation and guardrail enforcement as a separate, auditable layer.
  • Choose Bedrock-hosted OpenAI models when you want to consolidate AI spend, monitoring, and governance across multiple model providers (Anthropic, Meta, Mistral, OpenAI) under a single AWS account and billing structure.

Availability

  • Status: Generally Available (GA) as of June 11, 2026.
  • Regions: US East (N. Virginia) confirmed; the announcement notes this launch expands availability to "additional AWS Regions," implying prior availability in at least one other region.
  • Models available: GPT-5.5, GPT-5.4, GPT OSS Safeguard 120B, GPT OSS Safeguard 20B, GPT OSS 120B, and GPT OSS 20B are all listed in the Bedrock OpenAI model catalog.
  • API access: Available through the Amazon Bedrock Responses API with support for streaming, server-side tool calling, client-side tool calling, and projects.
  • Context window: 272K tokens for both GPT-5.4 and GPT-5.5.
  • Input modalities: Text and image inputs supported; output is text.
  • Pricing: Not specified in the announcement; pricing is expected to follow Bedrock's standard on-demand token-based pricing model—consult the Bedrock pricing page for current rates.
  • Documentation: Model cards available at https://docs.aws.amazon.com/bedrock/latest/userguide/model-cards-openai.html.

Tags

Servicesbedrock
Typeregion-expansionnew-model
Conceptsgenaillmagentic-aimultimodalcoding-assistant
Use Casesdeveloper-toolsenterprise
Providersopenai
GeographyAMERICAS

Related Resources

AI Radar AWS

AWS AI/ML news — curated, researched, explained

An automated intelligence platform that curates, researches, and analyzes AWS AI/ML/GenAI announcements daily. Every report is backed by real research — the system reads linked blog posts and documentation to provide accurate, in-depth analysis.

How Each Report Is Generated

  1. Collection — Daily monitoring of the AWS "What's New" RSS feed
  2. Filtering — AI-powered relevance detection for AI/ML/GenAI topics
  3. Taxonomy Tagging — LLM-based classification across 6 dimensions
  4. Importance Scoring — Point-based system with tag bonuses (1-5 stars)
  5. Research Phase — Follows links to blog posts and documentation
  6. Report Generation — Claude Sonnet produces structured 6-section analysis
  7. Visual Summary — Claude Opus generates Mermaid diagrams for key items
  8. Publishing — Static website rebuilt and deployed via CloudFront

Features

  • Faceted filtering by service, type, concept, and more
  • Multi-dimensional taxonomy with 80+ tags across 6 dimensions
  • Geographic availability badges (Global, APJ, EMEA, AMER) with filtering
  • Timeline visualization of announcement volume
  • PDF export for offline reading
  • Mermaid visual summaries for key announcements
  • Daily automated updates — no manual curation
What makes this different: Each report involves a dedicated research phase where the system reads linked blog posts and AWS documentation pages. This produces analysis that goes beyond the original announcement text.

Technology

Built with Python, AWS Lambda, Amazon Bedrock (Claude Sonnet 4.6, Opus 4.6, Haiku 4.5), S3, CloudFront, WAF, EventBridge, and CDK.

Open Source

This project is open source. Fork it, customize it for your needs, and deploy your own instance.
📦 github.com/bbonik/ai-radar-aws

How Importance Scoring Works

Each announcement receives a point score based on multiple factors. The total score maps to a 1-5 star rating:

1★ < 2 pts 2★ ≥ 2 pts 3★ ≥ 3.5 pts 4★ ≥ 5 pts 5★ ≥ 6.5 pts

Point Breakdown

FactorPointsWhen
Core AI service (Bedrock, AgentCore, SageMaker AI)+4Service named in title
Key AI service (SageMaker, Kiro, QuickSight)+2Service named in title
Other AI-related service+1Default
Blog post link+3Link to aws.amazon.com/blogs/
GitHub samples link+2Link to github.com/aws*
Documentation link+1Link to docs.aws.amazon.com/
New model+1.5Tagged as "new-model"
New service+1Tagged as "new-service"
New feature+0.5Tagged as "new-feature"
Anthropic / OpenAI provider+2Provider explicitly mentioned
Instance / notebook announcement-2Hardware/capacity, not feature
Performance / pricing / security-0.5Incremental updates
Region expansion to APJ+1Expands to Asia Pacific
Region expansion (non-APJ only)-1.5Only expands to other regions

Geographic Relevance Badges

Each announcement card shows a small badge indicating whether the feature is available in your region:

🌐 Global Available in all regions
🌏 APJ Asia Pacific
🌍 EMEA Europe / Middle East / Africa
🌎 AMER Americas (US, Canada, South America)
No badge Geography unknown
How geography is detected: The system detects ALL geographies mentioned in each announcement. If the text mentions specific regions (Tokyo, Frankfurt, Oregon, etc.), the corresponding geography badges are shown. If it says "all regions" or is a new feature with no region specified, it gets the Global badge. Geography is also filterable — click a geo chip to see only announcements available in that region.