
Bedrock Knowledge Base Cost Starts at the Vector Store
A 5,000-document corpus on the site calculator is $163.28/month on S3 Vectors and $510.52/month on OpenSearch Serverless Classic. $160 of the smaller bill is Claude Sonnet 5.5.
Tagged

A 5,000-document corpus on the site calculator is $163.28/month on S3 Vectors and $510.52/month on OpenSearch Serverless Classic. $160 of the smaller bill is Claude Sonnet 5.5.

AWS Settings spend limits, still a limited release on 23 Sep 2026, alert at 50%, 75%, and 90%, then pause the project. Do nothing for 90 days and AWS deletes the project data. The minimum limit is the greater of $20 or AWS's estimate.
On September 22, 2026 AWS made GPT-6 Sol and GPT-6 Luna generally available on Bedrock. OpenAI’s internal eval found Sol makes about half as many factual mistakes as GPT-5.6 Sol. On September 29 GPT-6.1 Sol followed. Ultrafast mode followed on October 8. Here is the lane routing.

On July 16, 2026 AWS removed the 30-day Standard residency gate for Standard-IA and One Zone-IA lifecycle transitions — day-0 IA can cut first-month storage ~46% on 10 TB cold logs, but the 30-day IA billing minimum and $0.01/GB retrieval fees still apply.

On June 15, 2026 Grok 4.3 GA on Bedrock Mantle at $1.25/$2.50 per 1M. On August 19, 2026 Grok 4.6 added Converse + US Geo/Global CRIS at $2.00–$2.20 / $6.00–$6.60. On September 28, 2026 Grok 4.7 shipped at the same rates — do not jump on name alone.

On July 30, 2026 AWS cut Bedrock on-demand prices 80% for GPT-5.6 Luna and 20% for Terra — Luna is now $0.22/$1.32 per 1M tokens. Here is the routing math and what not to re-price overnight.

On Jul 21, 2026 AWS shipped SES Essentials, Pro ($105/region), and Enterprise ($500/region). At 2M emails/mo a Pro-like a-la-carte stack runs ~$1,793 vs Pro plan ~$553 — switch only when you need the bundled stack.

Q2 2026 ranked: AgentCore Harness GA (Jun 17), FinOps Agent preview (Jun 9), Graviton5 GA (Jun 10), OpenAI on Bedrock (Apr 28). Adoption matrix + verified TCO signals for CTOs.

CloudZero, Vantage, Finout, nOps, ProsperOps, and Kubecost on AWS — platform selection guide plus who implements tagging, allocation, and architecture savings.

Production guide for Kubecost on AWS EKS — cost allocation setup plus architecture changes that reduce spend, not just attribute it.

Implement ProsperOps on AWS — Savings Plans automation works best after baseline modeling and architecture stability. Production checklist included.

EC2 On-Demand is a matrix of family, size, OS, tenancy, and region. In us-east-1 (June 2026) a m7g.large Linux runs $0.0816/hr vs a g5.xlarge GPU at $1.006/hr. Graviton5 M9g/C9g now GA for new fleets.