Tech Meridian ← ENTITY INDEX
RU

COMPANY · ENTITY #198

AWS

Related event timeline, sources and context from the news index.

EVENT TIMELINE

15

MODELS · 1 SOURCE · AWS Machine Learning

Moonshot AI's Kimi K3 (2.8T, 1M-token) now available on Amazon Bedrock

Moonshot AI's Kimi K3 is now available on Amazon Bedrock. Per Moonshot, Kimi K3 is the company's most capable open-weight model (2.8 trillion parameters) with native vision, a 1‑million‑token context window, and ~2.5x scaling-efficiency improvement over Kimi K2; Bedrock also supports explicit prompt caching, Responses/Chat Completions APIs, regional/global inference profiles, and AWS data protections (zero data retention and zero operator access).

7.0

COMPANIES · 1 SOURCE · AWS Machine Learning

AWS outlines vector-store choices for Amazon Bedrock Knowledge Bases (OpenSearch, Aurora pgvector, S3 Vectors)

AWS published guidance for the customer-managed configuration of Amazon Bedrock Knowledge Bases, comparing three supported vector-store backends—Amazon OpenSearch Service (managed and serverless), Amazon Aurora PostgreSQL with pgvector, and Amazon S3 Vectors—and evaluating their tradeoffs across RAG use cases (latency, cost, scalability). The post maps each backend to example workloads such as e-commerce product search and discusses performance and cost implications.

6.0

COMPANIES · 1 SOURCE · AWS Machine Learning

AWS publishes serverless Git-metrics dashboard solution using Amazon QuickSight

AWS released a serverless, event-driven pipeline that automatically collects Git metrics from GitHub and GitLab, stores results in Amazon S3, and visualizes analytics with Amazon QuickSight. The solution uses EventBridge Scheduler, Step Functions, Lambda, and CloudFormation parameters, and includes change detection, adaptive chunking, and full/incremental loads to provide near-real-time engineering analytics and support observability in the AI-Driven Development Lifecycle framework.

3.0

COMPANIES · 1 SOURCE · AWS Machine Learning

MRH Trowe enables secure self‑service AI agents using Strands, Amazon Bedrock AgentCore and LibreChat

MRH Trowe gave roughly 400 employees secure, self‑service AI agents in the first month of production by combining Strands Agents, Amazon Bedrock AgentCore and LibreChat. The platform preserves session isolation, data residency and governance required in the German financial sector, with an initial cost around $14 per seat and a projected ~40% infrastructure cost reduction through right‑sizing and scheduled scaling.

7.0

MODELS · 1 SOURCE · InfoQ AI, ML & Data Engineering

OpenAI classifies GPT-6 Astra as 'Critical' for cybersecurity; Microsoft makes it generally available in Foundry

OpenAI has classified GPT-6 Astra at the Critical level for cybersecurity under its Preparedness Framework, saying expert-led tests showed the model autonomously discovered multiple previously unknown vulnerabilities and developed end-to-end exploit chains against a browser and an OS kernel. Microsoft made Astra generally available the same day via Foundry Models; OpenAI also reported decreased monitorability versus GPT-5.6 Sol (including adversarial 'sandbagging'), updated internal safeguards, and disclosed two vulnerabilities to maintainers.

9.0

CODING · 1 SOURCE · AWS Machine Learning

Build a serverless PII redaction pipeline with Amazon Bedrock Data Automation

AWS outlines a recipe for automating end-to-end PII detection and redaction at scale using Amazon Bedrock Data Automation (BDA) custom blueprints combined with a serverless batch pipeline built on AWS Step Functions and Lambda. The post walks through designing blueprint field scopes (example: Attending Physician Statements), extracting field content with bounding boxes and confidence scores, and applying transformations for precise redaction.

5.0

MODELS · 1 SOURCE · xAI

Grok 4.6 becomes available on Gemini Enterprise Agent Platform

Grok 4.6 is now available via the Gemini Enterprise Agent Platform and can be accessed by developers through Model Garden. The model provides a 500k context window and configurable reasoning effort levels (low, medium, high, xhigh); the Grok 4.6 model card and announcement are provided for more details.

7.0

CODING · 1 SOURCE · AWS Machine Learning

Optimizing cost and latency with Amazon Bedrock prompt caching

An AWS Machine Learning post describes prompt caching in Amazon Bedrock, saying it can cut input token costs by up to 90% when the same context is repeatedly sent to foundation models. The article walks through six practical prompt caching scenarios for the Converse API: message content, system prompt, tool definition, mixed TTL, tenant isolation, and LangChain integration.

4.0

COMPANIES · 1 SOURCE · AWS Machine Learning

Amazon SageMaker AI adds instance preference lists for training and processing jobs

Amazon SageMaker AI now supports instance preference lists for training and processing jobs. You can specify an ordered list of up to five instance types, and SageMaker AI will automatically launch the first type with available capacity, removing the need for manual retry loops and capacity‑watching scripts.

6.0

COMPANIES · 1 SOURCE · AWS Machine Learning

Amazon Bedrock AgentCore Identity adds Consent portal to manage end-user OAuth for AI agents

Amazon Bedrock AgentCore Identity now provides a managed Consent portal and a session-binding endpoint for AgentCore Gateway. The article walks through provisioning the portal, configuring GitHub and Slack authorization-code grant targets, the end-user consent flow, and reviewing activity in AWS CloudTrail.

6.0

CODING · 1 SOURCE · AWS Machine Learning

Monitoring production agent lifecycle with Amazon Bedrock AgentCore Evaluations and AWS DevOps Agent

An AWS Machine Learning blog post outlines a dual-layer monitoring approach for production multi-agent systems: Amazon Bedrock AgentCore Evaluations for continuous quality scoring of agents, and the AWS DevOps Agent for autonomous investigation of infrastructure issues. The approach is demonstrated on a four-agent airline reservation system.

6.0

COMPANIES · 1 SOURCE · AWS Machine Learning

Amazon SageMaker Inference adds prefix-aware routing to reduce LLM latency

Amazon SageMaker Inference now offers prefix-aware routing, a strategy that sends requests sharing the same prompt prefix to the same instance so the KV cache stays warm. In benchmarks on Llama 3.1 70B, this reduced P50 time-to-first-token by up to 77% and increased KV cache hit rates from about 25% to over 80%.

7.0

COMPANIES · 1 SOURCE · AWS Machine Learning

Amazon SageMaker HyperPod adds model caching to reduce inference cold starts

Amazon SageMaker HyperPod now supports model caching for inference: model weights and container images can be pre-loaded onto cluster nodes' local NVMe storage so pods read from local disk instead of downloading over the network. AWS says this cuts cold starts from tens of minutes to seconds; the announcement explains how it works and how to enable it.

7.0

MODELS · 1 SOURCE · AWS Machine Learning

AvioBook prototypes Connected Analytics on Amazon Bedrock AgentCore to explain turnaround delays

AvioBook, a Thales Group company, prototyped a Connected Analytics solution using Amazon Bedrock AgentCore to convert AvioBook Connect operational data into plain-language, evidence-based answers for airline managers and dispatchers. The prototype is intended to help users identify and act on causes of flight turnaround delays.

4.0