AWS News - 2026-08-28
2026-08-28
最終更新: 2026-08-29 05:39:36 JST
AI による概要
この日はインド国内での推論とコスト最適化、そしてモデルの品揃え拡充が話題でした。Amazon Bedrock が OpenAI GPT-5.6 の Terra と Luna をインド地理のクロスリージョン推論でサポートし、ローカルのデータ保護要件があるケースに対応します。SageMaker JumpStart には NVIDIA の Cosmos3-Edge / Cosmos3-Nano / Cosmos3-Super、Meta の Muse-Glimmer-30B、Alibaba の Qwen 3.8-27B が追加され、AWS GovCloud (US) の Bedrock では長時間の推論やエージェント作業を狙った SpaceXAI の Grok 4.6 が使えるようになりました。機械学習ブログでは NVIDIA CUDA Multi-Process Service を使い、1 リクエストが GPU の一部しか使わない音声認識の推論コストを 75% 削減する手法が公開されています。セキュリティでは Bedrock Guardrails をモデル境界だけでなくツール呼び出しにも広げる方法が Strands Agents SDK を使って解説されました。サービス更新では Elastic Disaster Recovery に複数サーバーのアプリケーションを順序立てて起動する Recovery Plans が追加され、AgentCore が北カリフォルニアとハイデラバードの 2 リージョンへ拡大、FSx for NetApp ONTAP はリージョン間・アカウント間のバックアップコピーに対応しています。国内では Bedrock のきめ細かなコスト配分と Athena / CUDOS を使った可視化の日本語解説、8 月 25 日に 20 周年を迎えた Amazon EC2 の振り返りに加え、Kiro IDE の diagnostics ツールが返す静的解析の指摘を手がかりに、AI コーディングエージェントが本当に良くなっているのかを検証する記事も公開されました。
主要トピック
データレジデンシー: Bedrock が GPT-5.6 Terra / Luna をインド地理のクロスリージョン推論でサポート
モデル拡充: SageMaker JumpStart に NVIDIA Cosmos3 ファミリー、Meta Muse-Glimmer-30B、Alibaba Qwen 3.8-27B を追加
モデル拡充: AWS GovCloud (US) の Bedrock で SpaceXAI Grok 4.6 が利用可能に
コスト最適化: NVIDIA CUDA MPS により音声認識 (ASR) の推論コストを 75% 削減
コスト可視化: Bedrock のきめ細かなコスト配分と、Athena / CUDOS を使った分析の日本語解説
ガードレール: Bedrock Guardrails を Strands Agents SDK でツール呼び出しにも拡張
災害対策: Elastic Disaster Recovery の Recovery Plans で複数サーバーアプリの順序立てた起動を自動化
バックアップ: FSx for NetApp ONTAP がリージョン間・アカウント間のバックアップコピーに対応
リージョン拡大: AgentCore が米国西部 (北カリフォルニア) とアジアパシフィック (ハイデラバード) へ拡大
開発ツール: Kiro IDE の diagnostics による静的解析の指摘から、AI コーディングエージェントの改善度を検証
国内: 20 周年を迎えた Amazon EC2 の振り返り、AWS Local Executive Roadshow 札幌編の開催レポート 2 本
AWS What's New
Amazon EVS now supports i7i.metal-48xl Amazon EC2 instance type
- Link: https://aws.amazon.com/about-aws/whats-new/2026/08/amazon-evs-i7i-48xl
- Published: 2026-08-28 01:35:00
- Fetched: 2026-08-28 08:31:52
詳細を表示
Today, we're announcing that Amazon Elastic VMware Service (Amazon EVS) now supports the i7i.metal-48xl Amazon Elastic Cloud Compute (Amazon EC2) bare-metal instance type, offering a higher core-count option with a newer generation processor to help you realize cost-performance benefits for your VMware-based workloads on AWS.
With this release, you now have more options for running your virtual machines (VMs) on Amazon EVS environments and growing your cloud presence at your own pace, as your business demands. Powered by 5th generation Intel Xeon Scalable processors, i7i instances offer the best compute and storage performance for x86-based storage optimized instances in Amazon EC2, delivering up to 23% better compute performance and more than 10% better price performance over i4i instances. The i7i.metal-48xl's larger memory and storage footprint allows you to run more VMs to vertically scale your operation with less EVS hosts - and take further advantage of the latest VMware Cloud Foundation (VCF) 9.x features such as memory tiering and new automations such as the Amazon EVS Deployment Orchestrator to simplify onboarding.
This latest release is available in AWS Regions where Amazon EVS and Amazon EC2 i7i are both available. See Amazon EVS regional availability and Amazon EC2 i7i regional availability.
Learn more about Amazon EVS by visiting the product detail page and the user guide.
AWS Backup adds cross-Region and cross-account backup support for Amazon FSx for NetApp ONTAP
- Link: https://aws.amazon.com/about-aws/whats-new/2026/08/aws-backup-amazon-fsx-netapp-cross-account-region/
- Published: 2026-08-28 02:25:00
- Fetched: 2026-08-28 08:31:52
AWS Backup now supports copying Amazon FSx for NetApp ONTAP backups across AWS Regions and accounts. You can copy your FSx for NetApp ONTAP backups to another AWS Region, another AWS account, or both, using policy-based backup plans or on-demand copy jobs.
Copying backups to a separate Region helps you meet cross-Region disaster recovery and business continuity goals, while copying to a separate account helps protect your backups from accidental deletion, operational error, or account compromise. AWS Backup manages these copies centrally, and you can automate them across your organization using AWS Organizations.
This capability is available in all commercial AWS Regions where both Amazon FSx for NetApp ONTAP and AWS Backup are available. To get started, visit the AWS Backup console, AWS Command Line Interface (CLI), or AWS SDKs. For a complete list of supported Regions and features, visit the AWS Backup documentation. To learn more, visit the product page and pricing page.
Amazon FSx for NetApp ONTAP now supports copying backups across AWS Regions and accounts
- Link: https://aws.amazon.com/about-aws/whats-new/2026/08/fsx-ontap-cross-region-backup-copy/
- Published: 2026-08-28 02:37:00
- Fetched: 2026-08-28 08:31:52
Amazon FSx for NetApp ONTAP, a fully managed shared storage service built on NetApp's popular ONTAP file system, now supports copying backups within and across AWS Regions, and across trusted accounts in your AWS organization. You can now provide an additional layer of protection for your data by storing secondary backup copies in a different Region or account than your primary copies, making it easier to meet your business continuity, data protection, and compliance requirements.
Amazon FSx offers secure, highly durable, and incremental backups designed to support your data retention, archival, and compliance needs. Backups are point-in-time offline copies of your volumes, stored redundantly across multiple Availability Zones in the same Region as your file system. Previously, you could create and restore backups in the same Region and account as your file system. Starting today, you can copy backup data to multiple Regions and accounts for enhanced resilience, data isolation, and operational flexibility.
You can copy new and existing backups in all AWS Regions where Amazon FSx for NetApp ONTAP is available. To learn more, visit the FSx for NetApp ONTAP user guide and the Amazon FSx for NetApp ONTAP product page.
Amazon Connect Customer now automatically refreshes scheduling metrics
- Link: https://aws.amazon.com/about-aws/whats-new/2026/08/amazon-connect-customer-scheduling-metrics/
- Published: 2026-08-28 02:59:00
- Fetched: 2026-08-28 08:31:52
Amazon Connect Customer now automatically refreshes schedule metrics on the scheduling page, giving managers immediate visibility into the impact of schedule updates. For example, when a new team meeting is added for 10 agents from 10 AM – 11 AM, metrics such as net available headcount and projected service level are automatically updated. This gives workforce managers up-to-date visibility into how schedule changes affect staffing coverage, so they can make faster, more informed decisions throughout the day.
This feature is available in all AWS Regions where Amazon Connect Customer agent scheduling is available. To learn more about Amazon Connect Customer agent scheduling, click here .
AWS Elastic Disaster Recovery introduces Recovery Plans for orchestrated application recovery
- Link: https://aws.amazon.com/about-aws/whats-new/2026/08/elastic-disaster-recovery-plans/
- Published: 2026-08-28 03:00:00
- Fetched: 2026-08-28 08:31:52
AWS Elastic Disaster Recovery (AWS DRS) now offers Recovery Plans, a capability that automates the sequential launch of multi-server applications during recovery and drills. Instead of launching servers one at a time and tracking dependencies manually, you define the recovery sequence once and run it with a single action when you need it. This addresses a common challenge for customers running applications made up of databases, application tiers, and supporting services that must come up in a specific order.
Recovery Plans let you group servers into sequential steps with configurable wait times between them, so your application recovers predictably and consistently every time. You can run plans in a non-disruptive drill mode to validate your procedures, add approval steps where you want a person in the loop, and monitor progress in real time. This automation reduces recovery time and eliminates common errors that occur when coordinating server launch order during high-stress disaster scenarios.
Recovery Plans are available in all AWS Regions where AWS DRS is offered, at no additional cost beyond standard DRS usage. To learn more, visit the AWS Elastic Disaster Recovery User Guide.
Amazon Bedrock AgentCore expands to two new regions
- Link: https://aws.amazon.com/about-aws/whats-new/2026/08/bedrock-agentcore-two-new-regions/
- Published: 2026-08-28 03:00:00
- Fetched: 2026-08-28 08:31:52
Amazon Bedrock AgentCore is now available in two additional AWS Regions: US West (N. California) and Asia Pacific (Hyderabad). Amazon Bedrock AgentCore is the platform to build, connect, and optimize agents. It helps engineers ship agents fast with any framework and any model, connect them to enterprise systems and tools, and optimize them continuously, with security enforced at the infrastructure layer that agents can't bypass.
With this expansion, customers in these regions can build and run agents closer to their end users with lower latency. AgentCore capabilities including agent runtime, identity and access control, policy management, session persistence, tool connectivity, evaluations and observability are available in these regions at launch.
For more information on AgentCore, visit the AgentCore product page or the AgentCore Developer Guide. To learn about pricing, visit AgentCore pricing. For region availability, visit Supported AWS Regions.
Amazon Connect Customer expands conversational analytics capabilities in the Africa (Cape Town) Region
- Link: https://aws.amazon.com/about-aws/whats-new/2026/08/connect-customer-analytics-cape-town/
- Published: 2026-08-28 03:00:00
- Fetched: 2026-08-28 08:31:52
Amazon Connect Customer now supports generative AI-powered summaries, real-time call analytics, and real-time rules in the Africa (Cape Town) Region. These capabilities extend the post-contact conversational analytics already available in the Region, giving companies access to AI-powered insights that help them improve customer experience, agent performance, and operational efficiency.
With generative AI-powered summaries, supervisors and human-agents receive concise, AI-generated summaries of customer interactions immediately after a contact ends, eliminating the need for manual note-taking and reducing after-contact work. Real-time call analytics provide live transcription enabling supervisors to monitor interactions as they happen and intervene when needed to improve outcomes. Real-time rules automatically categorize contacts and alert supervisors based on keywords, sentiment, and other criteria detected during a live interaction, enabling immediate action such as sending notifications or generating tasks.
To learn more about Connect Customer conversational analytics capabilities, refer to the Amazon Connect Customer Administrator Guide or visit the Amazon Connect Customer website. For a complete list of conversational analytics features available by Region, refer to Availability of Connect Customer features by Region.
Amazon EC2 X8i instances are now available in additional regions
- Link: https://aws.amazon.com/about-aws/whats-new/2026/08/amazon-ec2-x8i-europe-milan-spain/
- Published: 2026-08-28 03:40:00
- Fetched: 2026-08-28 08:31:52
Starting today, Amazon Elastic Compute Cloud (Amazon EC2) X8i instances are available in the Europe (Milan) and Europe (Spain) regions. These instances are powered by custom Intel Xeon 6 processors available only on AWS. X8i instances are SAP-certified and deliver the highest performance and fastest memory bandwidth among comparable Intel processors in the cloud. They deliver up to 43% higher performance, 1.5x more memory capacity (up to 6TB), and 3.3x more memory bandwidth compared to previous generation X2i instances.
X8i instances are designed for memory-intensive workloads like SAP HANA, large databases, data analytics, and Electronic Design Automation (EDA). Compared to X2i instances, X8i instances offer up to 50% higher SAPS performance, up to 47% faster PostgreSQL performance, 88% faster Memcached performance, and 46% faster AI inference performance. X8i instances come in 14 sizes, from large to 96xlarge, including two bare metal options.
To get started, visit the AWS Management Console. X8i instances can be purchased via Savings Plans, On-Demand instances, and Spot instances. For more information visit X8i instances page.
Amazon Redshift integrates with Agent Toolkit for AWS for AI-assisted data warehouse management
- Link: https://aws.amazon.com/about-aws/whats-new/2026/08/redshift-agenttoolkit-for-ai-assisted-datawarehouse-mgmt
- Published: 2026-08-28 05:07:00
- Fetched: 2026-08-28 08:31:52
詳細を表示
Amazon Redshift now integrates with the Agent Toolkit for AWS, enabling you to build, query, troubleshoot, and migrate to Amazon Redshift data warehouses and date lakes directly from AI agents such as Claude Code, Kiro, and Cursor. The integration pairs the AWS MCP (Model Context Protocol) server, which provides authenticated AWS API execution on your behalf — with Redshift skills: curated packages of tested procedures and reference material that help AI agents complete Redshift tasks more effectively.
The Redshift skills cover SQL syntax references to reduce query generation errors, metadata discovery to explore schemas and data without writing SQL by hand, data loading patterns, materialized view best practices, function and data type guidance, and extensions such as Qualify, Pivot, and Super. It also guides end-to-end data warehouse migrations to Amazon Redshift, including discovery, schema and SQL conversion, data movement, validation, and performance comparison. We will continue to expand these skills with additional capabilities over time.
The skills work with provisioned clusters and Serverless workgroups, require no changes to existing infrastructure, and are available at no additional charge in all AWS Regions where Amazon Redshift and the AWS MCP Server are offered.
To get started, install the aws-data-analytics plugin in your agent, which bundles the MCP Server configuration and Redshift skills in a single step. Agents with MCP Server access can also discover and load skills at runtime without pre-installation. For setup instructions, see the Agent Toolkit documentation or the Amazon Redshift skills documentation.
Amazon Redshift streaming can now ingest 10MiB records from Amazon Kinesis Data Streams
- Link: https://aws.amazon.com/about-aws/whats-new/2026/08/redshift-streaming-supports-kds-10mib-records
- Published: 2026-08-28 06:41:00
- Fetched: 2026-08-28 08:31:52
Amazon Redshift now supports Amazon Kinesis Data Streams (KDS) record sizes up to 10 MiB—a 10x increase from the previous 1 MiB limit—fully matching the expanded maximum record size in Amazon KDS. This means you can stream significantly larger payloads directly into Amazon Redshift without splitting records, simplifying your ingestion pipelines and unlocking new use cases for high-volume, large-record workloads.
Amazon Redshift support for 10MiB record size in Amazon KDS streams is now available in all commercial AWS regions where Amazon Redshift is available. For more information on direct streaming ingestion into Amazon Redshift, see the Amazon Redshift streaming documentation. For more information on 10MiB record support in Amazon KDS, see the Amazon KDS documentation.
Muse-Glimmer-30B and Qwen 3.8-27B models now available on Amazon SageMaker JumpStart
- Link: https://aws.amazon.com/about-aws/whats-new/2026/01/muse-glimmer-30b-qwen-3.8-27b-on-sagemaker-jumpstart/
- Published: 2026-08-28 07:39:00
- Fetched: 2026-08-28 12:11:27
詳細を表示
Meta's Muse-Glimmer-30B and Alibaba's Qwen 3.8-27B models are now available on Amazon SageMaker JumpStart, expanding the portfolio of foundation models available to AWS customers. These two models bring specialized capabilities spanning autonomous local agentic workflows and multimodal long-horizon reasoning, enabling customers to deploy high-performance, scalable AI solutions on AWS infrastructure.
These models address different enterprise AI challenges with specialized capabilities:
Muse-Glimmer-30B is engineered for autonomous agentic tasks with multi-step reasoning, tool use, and failure recovery. This 30B-parameter dense model from Meta Superintelligence Lab combines a dedicated ~1.8B ViT-G/14 perception encoder with interleaved text and image inputs, a 131K+ context window, and selectable reasoning strength (low through extra-high). Released under Apache 2.0, it handles sequential tool calls, recovers from failures, and operates entirely without cloud infrastructure which is ideal for always-on enterprise agents.
Qwen 3.8-27B excels in coding, multi-step agentic tasks, and multimodal understanding across text, images, and video. A dense 27B-parameter native vision-language model with a 262K context window (extendable to ~1M via YaRN scaling), it delivers substantial gains over its predecessor with adjustable reasoning effort levels. Scoring 61.7 on SWE-bench Pro and running at ~17GB quantized, it carries complex multi-step tasks through to completion with greater reliability.
With SageMaker JumpStart, customers can deploy any of these models with just a few clicks to address their specific AI use cases.
To get started with these models, navigate to the SageMaker JumpStart model catalog in the SageMaker console or use the SageMaker Python SDK to deploy the models to your AWS account. For more information about deploying and using foundation models in SageMaker JumpStart, see the Amazon SageMaker JumpStart documentation.
Cosmos3-Edge, Cosmos3-Nano, and Cosmos3-Super models now available on Amazon SageMaker JumpStart
- Link: https://aws.amazon.com/about-aws/whats-new/2026/01/cosmos3-edge-cosmos3-nano-cosmos3-super-on-sagemaker-jumpstart/
- Published: 2026-08-28 07:41:00
- Fetched: 2026-08-28 12:11:27
詳細を表示
NVIDIA's Cosmos3-Edge, Cosmos3-Nano, and Cosmos3-Super models are now available on Amazon SageMaker JumpStart, expanding the portfolio of foundation models available to AWS customers. These three models form the Cosmos 3 family of open, frontier omnimodal world models for physical AI, enabling customers to build robots, autonomous vehicles, and vision AI that perceive, reason, plan, and act in the physical world.
These models address different physical AI challenges with specialized capabilities:
Cosmos3-Edge is engineered for on-device robot control and real-time visual reasoning on edge hardware. This 4B-parameter omni-model (with a 2B Nemotron-based reasoner) operates at robot-control resolution (640×360), delivering real-time reasoning and generating 32 actions per inference at 15 Hz on NVIDIA Jetson Thor. It supports 256p and 480p video at 12–30 FPS, bringing frontier physical AI capabilities directly to embedded systems.
Cosmos3-Nano excels in physics-aware world generation and physical reasoning as a compact 16B-parameter omnimodal model. It processes combinations of text, image, video, audio, and action trajectories to produce corresponding outputs, enabling robots and vision AI agents to reason using prior knowledge, physics understanding, and common sense. It supports chain-of-thought reasoning over text, images, and video with resolutions up to 720p.
Cosmos3-Super provides the highest-fidelity world generation and simulation in the Cosmos 3 family at 64B parameters. It jointly processes and generates language, images, video, audio, and action sequences within a unified Mixture-of-Transformers architecture, supporting resolutions up to 720p across multiple aspect ratios. Ideal for large-scale simulation, synthetic data generation, and policy learning workflows.
With SageMaker JumpStart, customers can deploy any of these models with just a few clicks to address their specific AI use cases.
To get started with these models, navigate to the SageMaker JumpStart model catalog in the SageMaker console or use the SageMaker Python SDK to deploy the models to your AWS account. For more information about deploying and using foundation models in SageMaker JumpStart, see the Amazon SageMaker JumpStart documentation.
SpaceXAI Grok 4.6 now available on Amazon Bedrock in AWS GovCloud (US)
- Link: https://aws.amazon.com/about-aws/whats-new/2026/08/spacexai-grok-4-6-govcloud/
- Published: 2026-08-28 22:17:00
- Fetched: 2026-08-29 05:39:35
Amazon Bedrock in AWS GovCloud (US) now supports SpaceXAI Grok 4.6, a frontier model built for coding, agentic tasks, and knowledge work. Grok 4.6 is SpaceXAI’s latest flagship model, built for long-running agents and ambitious interactive and visual work. It offers 500k context window and configurable reasoning efforts (low, medium, high, xhigh).
The model runs on the bedrock-runtime endpoint with support for the Responses, Chat Completions, and Converse APIs, and customers can access Grok 4.6 at scale with cross-Region inference routing requests across both AWS GovCloud (US) Regions. Grok 4.6 is also avialable via the bedrock-mantle endpoint in AWS GovCloud (US-East).
To get started, review the model card for Grok 4.6 in the Amazon Bedrock User Guide.
AWS Japan Blog
Amazon EC2、20 回目の誕生日おめでとう
- Link: https://aws.amazon.com/jp/blogs/news/happy-20th-birthday-amazon-ec2/
- Published: 2026-08-28 11:39:40
- Fetched: 2026-08-28 12:11:28
実践企業に学ぶ生成 AI 導入の勘所 〜眠るデータを企業価値に変える〜 – AWS Local Executive Roadshow 札幌編(#7/8)開催レポート
- Link: https://aws.amazon.com/jp/blogs/news/local_executive_roadshow_7/
- Published: 2026-08-28 11:45:20
- Fetched: 2026-08-28 12:11:28
AI ツールで実現する継続収益ビジネス 〜開発力を資産に変える〜 – AWS Local Executive Roadshow 札幌編(#8/8)開催レポート
- Link: https://aws.amazon.com/jp/blogs/news/local_executive_roadshow_8/
- Published: 2026-08-28 13:08:54
- Fetched: 2026-08-28 17:50:05
パート2: Amazon AthenaとCUDOSを使用したAmazon Bedrockのコスト配分
- Link: https://aws.amazon.com/jp/blogs/news/part-2-amazon-bedrock-cost-attribution-with-amazon-athena-and-cudos/
- Published: 2026-08-28 16:24:33
- Fetched: 2026-08-28 17:50:05
Amazon Bedrock のきめ細かなコスト配分の導入
- Link: https://aws.amazon.com/jp/blogs/news/introducing-granular-cost-attribution-for-amazon-bedrock/
- Published: 2026-08-28 16:40:20
- Fetched: 2026-08-28 17:50:05
AI コーディングエージェントは本当に良くなっているのか?
- Link: https://aws.amazon.com/jp/blogs/news/diagnostics-over-time/
- Published: 2026-08-28 20:59:49
- Fetched: 2026-08-29 05:39:36
AWS Security Blog
Extend Amazon Bedrock Guardrails to Tool Interactions Using the Strands Agents SDK
- Link: https://aws.amazon.com/blogs/security/extend-amazon-bedrock-guardrails-to-tool-interactions-using-the-strands-agents-sdk/
- Published: 2026-08-28 01:20:05
- Fetched: 2026-08-28 08:31:53
AWS Machine Learning Blog
Reduce ASR inference costs by 75% with NVIDIA MPS on Amazon EC2
- Link: https://aws.amazon.com/blogs/machine-learning/reduce-asr-inference-costs-by-75-with-nvidia-mps-on-amazon-ec2/
- Published: 2026-08-28 01:05:10
- Fetched: 2026-08-28 08:31:54
Deepgram deepens Amazon SageMaker AI observability with Enhanced Metrics
- Link: https://aws.amazon.com/blogs/machine-learning/deepgram-deepens-amazon-sagemaker-ai-observability-with-enhanced-metrics/
- Published: 2026-08-28 01:11:27
- Fetched: 2026-08-28 08:31:54
Introducing OpenAI models on Amazon Bedrock for in-country inferencing in India
- Link: https://aws.amazon.com/blogs/machine-learning/introducing-openai-models-on-amazon-bedrock-for-in-country-inferencing-in-india/
- Published: 2026-08-28 03:36:08
- Fetched: 2026-08-28 08:31:54
Build agentic creative workflows with Amazon Quick and fal
- Link: https://aws.amazon.com/blogs/machine-learning/build-agentic-creative-workflows-with-amazon-quick-and-fal/
- Published: 2026-08-28 08:04:22
- Fetched: 2026-08-28 08:31:54