Trending Posts

Reduce LLM latency with prefix-aware routing on Amazon SageMaker Inference

When you build an application on top of a large language model (LLM), the prompt…

ByByadmin Sep 11, 2026 12 min read

Reduce inference cold starts on Amazon SageMaker HyperPod with model caching

When you deploy a large language model (LLM) for inference on Amazon SageMaker HyperPod, there’s…

ByByadmin Sep 11, 2026 12 min read

Video and image search in Amazon Bedrock Knowledge Base using Marengo 3.0

Today we’re announcing the general availability of TwelveLabs Marengo Embed 3.0 as an embedding model…

ByByadmin Sep 11, 2026 8 min read

Amazon Quick is now generally available on desktop

Your teams get an AI assistant that handles real work while your data stays in…

ByByadmin Sep 11, 2026 7 min read

Build an end-to-end RFI questionnaire workflow using Amazon Quick Automate

Build an end-to-end Request for Information (RFI) questionnaire workflow using Amazon Quick Automate to solve…

ByByadmin Sep 11, 2026 12 min read

Model-agnostic PII detection with LLMs

A configurable, instruction-driven detector that runs on any large language model (LLM) managed on Amazon…

ByByadmin Sep 11, 2026 19 min read

3 ways to prep for your next big race with Search

Search can help runners get race-day ready with registration alerts, tailored training plans, and more.…

ByByadmin Sep 11, 2026 1 min read

Agent Evaluation Metric for multi-turn conversations

Multi-turn agents fail in ways that single-turn evaluation misses: one early mistake quietly corrupts every…

ByByadmin Sep 10, 2026 21 min read

How AvioBook builds turnaround insights from operational data with Amazon Bedrock AgentCore

This post is co-written with Petra Lafond, Product Manager, and Maarten Cardinaels, Tech Lead at…

ByByadmin Sep 10, 2026 15 min read
Scroll to Top