← Back to all announcements
★★☆☆☆ 27/05/2026

AWS Elemental Inference now supports Smart Subtitles for automated live captioning

Broadcasters can now auto-caption live streams in 6 languages with no third-party tools—and save more by bundling it with other AI video features.

View original announcement →

Visual Summary

graph TD A{{Smart Subtitles}}:::announced B(AWS Elemental MediaLive):::compute C(AWS Elemental Inference):::compute D([Custom Dictionaries]):::feature E([ASR Transcription]):::feature F([TTML/WebVTT Output]):::feature G(MediaPackage):::compute H((Broadcaster)):::external I([Smart Cropping]):::feature H ==>|"configures"| B B ==>|"enables"| A A -->|"powered by"| C C -->|"applies"| E A -->|"generates"| F D -.->|"improves accuracy"| E F -->|"delivers to"| G A -.->|"bundles with"| I classDef announced fill:#ff9900,stroke:#ec7211,color:#fff,font-weight:bold classDef compute fill:#e3f2fd,stroke:#1565c0,color:#1565c0 classDef storage fill:#e8f5e9,stroke:#2e7d32,color:#2e7d32 classDef feature fill:#fff3e0,stroke:#e65100,color:#e65100 classDef external fill:#f5f5f5,stroke:#616161,color:#616161

What's New

AWS Elemental Inference has added Smart Subtitles, an AI-powered feature that automatically generates real-time subtitles for live video streams using automatic speech recognition (ASR). The feature delivers subtitles in TTML or WebVTT format with low latency and integrates natively with AWS Elemental MediaLive, eliminating the need for manual captioning workflows or third-party services. It supports eight language variants across six languages and includes custom dictionary support for domain-specific terminology.

How It Works

  • ASR-based transcription: Smart Subtitles applies automatic speech recognition to the audio track of live video streams, converting spoken content into timed subtitle text in real time.
  • Dual output formats: Subtitles are delivered in TTML format for MediaPackage V2, CMAF Ingest, and Microsoft Smooth output groups, and in WebVTT format for HLS and MediaPackage output groups.
  • Native MediaLive integration: The feature is enabled directly within an AWS Elemental MediaLive channel configuration via the console or API, requiring no external pipeline or third-party service.
  • Custom dictionaries: Users can create custom dictionaries through the Elemental Inference API or console to improve transcription accuracy for specialized vocabulary such as athlete names, technical jargon, or brand terms.
  • Multi-feature processing: Smart Subtitles shares underlying video/audio analysis with other Elemental Inference features (e.g., smart cropping, clip generation), enabling cost-efficient multi-feature deployments on the same stream.
  • Non-linear pricing: When multiple features are enabled simultaneously, a bundled per-minute rate applies (e.g., $0.23/min for two features vs. $0.30/min if priced separately), reducing effective per-feature cost.

Why It's Important

  • Accessibility compliance: Broadcasters can meet accessibility requirements (e.g., FCC captioning mandates, WCAG standards) for live content without building or procuring separate captioning infrastructure.
  • Operational simplification: Eliminating manual captioning workflows and third-party integrations reduces operational complexity, vendor management overhead, and potential points of failure in live production pipelines.
  • Broad language coverage: Support for English (US, GB, AU), French, German, Italian, Portuguese, and Spanish enables international broadcasters to serve multilingual audiences from a single integrated service.
  • Low-latency delivery: Real-time subtitle generation with low latency ensures subtitles remain synchronized with live content, which is critical for sports, news, and event broadcasting.
  • Cost efficiency at scale: The non-linear pricing model means organizations running multiple AI features simultaneously—such as vertical video cropping plus subtitles for an 8-hour sports day—pay significantly less than they would with separate point solutions.

How It's Different

  • Integrated vs. bolt-on: Unlike third-party captioning services that require separate API integrations, webhook pipelines, or manual operator intervention, Smart Subtitles is natively embedded in the MediaLive channel workflow.
  • Bundled AI pricing: Competing solutions typically charge independently per feature or per service; Elemental Inference's non-linear pricing reduces per-feature costs when combining Smart Subtitles with smart cropping or clip generation on the same content.
  • Custom dictionary support: Many automated captioning services offer limited or no domain-specific vocabulary customization; Elemental Inference allows custom dictionaries via API or console for specialized content verticals.
  • Multi-format output: The service outputs both TTML and WebVTT natively, covering the full range of common streaming delivery formats (HLS, CMAF, MediaPackage, Smooth Streaming) without additional transcoding steps.
  • Unified ML platform: Smart Subtitles is part of a broader real-time ML video service (Elemental Inference) that also handles smart cropping and clip generation, allowing teams to consolidate AI video capabilities under a single service rather than managing multiple vendors.

When to Prefer It

  • Live sports broadcasting: When commentary includes rapidly changing athlete names, team names, or play-by-play terminology that benefits from custom dictionaries to maintain transcription accuracy.
  • News and current affairs: When real-time captioning is required for regulatory compliance and latency must be minimized to keep subtitles synchronized with breaking news coverage.
  • Multilingual international streams: When content needs to be subtitled in French, German, Italian, Portuguese, or Spanish for regional audiences without deploying separate per-language captioning services.
  • Multi-feature AI workflows: When a broadcaster is already using smart cropping for vertical video or clip generation, adding Smart Subtitles at the bundled rate delivers meaningful cost savings versus a standalone captioning solution.
  • Organizations replacing third-party captioning vendors: When teams want to reduce vendor sprawl, simplify billing, and consolidate live AI video capabilities within the AWS ecosystem.
  • OTT platforms with HLS or CMAF delivery: When the delivery stack uses HLS (WebVTT) or CMAF/MediaPackage V2 (TTML) output groups and subtitle format compatibility is a key requirement.

Availability

  • General Availability: Smart Subtitles launched as a generally available feature on May 27, 2026.
  • Supported AWS Regions: Available in US East (N. Virginia), US West (Oregon), Asia Pacific (Mumbai), and Europe (Ireland).
  • Supported languages: English (United States, Great Britain, Australia), French, German, Italian, Portuguese, and Spanish.
  • Pricing model: Consumption-based, pay-as-you-go; single-feature rate is $0.15/min per pipeline; two-feature bundled rate is $0.23/min per pipeline in US East (N. Virginia).
  • Dual-pipeline note: When used with a Standard (dual-pipeline) MediaLive channel, costs are applied per pipeline, effectively doubling the per-minute charge for redundancy configurations.
  • Integration requirement: Requires AWS Elemental MediaLive; Smart Subtitles is enabled through the MediaLive channel configuration and managed via the Elemental Inference API or console.

Tags

Servicesother-aws
Typenew-featurega-launch
Conceptsspeechnlp
Use Casesenterprise
GeographyGlobal

Related Resources

AI Radar AWS

AWS AI/ML news — curated, researched, explained

An automated intelligence platform that curates, researches, and analyzes AWS AI/ML/GenAI announcements daily. Every report is backed by real research — the system reads linked blog posts and documentation to provide accurate, in-depth analysis.

How Each Report Is Generated

  1. Collection — Daily monitoring of the AWS "What's New" RSS feed
  2. Filtering — AI-powered relevance detection for AI/ML/GenAI topics
  3. Taxonomy Tagging — LLM-based classification across 6 dimensions
  4. Importance Scoring — Point-based system with tag bonuses (1-5 stars)
  5. Research Phase — Follows links to blog posts and documentation
  6. Report Generation — Claude Sonnet produces structured 6-section analysis
  7. Visual Summary — Claude Opus generates Mermaid diagrams for key items
  8. Publishing — Static website rebuilt and deployed via CloudFront

Features

  • Faceted filtering by service, type, concept, and more
  • Multi-dimensional taxonomy with 80+ tags across 6 dimensions
  • Geographic availability badges (Global, APJ, EMEA, AMER) with filtering
  • Timeline visualization of announcement volume
  • PDF export for offline reading
  • Mermaid visual summaries for key announcements
  • Daily automated updates — no manual curation
What makes this different: Each report involves a dedicated research phase where the system reads linked blog posts and AWS documentation pages. This produces analysis that goes beyond the original announcement text.

Technology

Built with Python, AWS Lambda, Amazon Bedrock (Claude Sonnet 4.6, Opus 4.6, Haiku 4.5), S3, CloudFront, WAF, EventBridge, and CDK.

Open Source

This project is open source. Fork it, customize it for your needs, and deploy your own instance.
📦 github.com/bbonik/ai-radar-aws

How Importance Scoring Works

Each announcement receives a point score based on multiple factors. The total score maps to a 1-5 star rating:

1★ < 2 pts 2★ ≥ 2 pts 3★ ≥ 3.5 pts 4★ ≥ 5 pts 5★ ≥ 6.5 pts

Point Breakdown

FactorPointsWhen
Core AI service (Bedrock, AgentCore, SageMaker AI)+4Service named in title
Key AI service (SageMaker, Kiro, QuickSight)+2Service named in title
Other AI-related service+1Default
Blog post link+3Link to aws.amazon.com/blogs/
GitHub samples link+2Link to github.com/aws*
Documentation link+1Link to docs.aws.amazon.com/
New model+1.5Tagged as "new-model"
New service+1Tagged as "new-service"
New feature+0.5Tagged as "new-feature"
Anthropic / OpenAI provider+2Provider explicitly mentioned
Instance / notebook announcement-2Hardware/capacity, not feature
Performance / pricing / security-0.5Incremental updates
Region expansion to APJ+1Expands to Asia Pacific
Region expansion (non-APJ only)-1.5Only expands to other regions

Geographic Relevance Badges

Each announcement card shows a small badge indicating whether the feature is available in your region:

🌐 Global Available in all regions
🌏 APJ Asia Pacific
🌍 EMEA Europe / Middle East / Africa
🌎 AMER Americas (US, Canada, South America)
No badge Geography unknown
How geography is detected: The system detects ALL geographies mentioned in each announcement. If the text mentions specific regions (Tokyo, Frankfurt, Oregon, etc.), the corresponding geography badges are shown. If it says "all regions" or is a new feature with no region specified, it gets the Global badge. Geography is also filterable — click a geo chip to see only announcements available in that region.