AWS Machine Learning Blog, read from an operations desk
Everything we have published under AWS Machine Learning Blog, read from an operations desk: what it changes for a B2B company of 10-200 people.
AWS Adds Access Controls for AI Agents Calling External Tools
If your company has connected an AI agent to your CRM, ticketing system, or internal APIs to automate sales outreach or support triage, that agent likely has more access than it needs and no audit trail of what it actually did. AgentCore Gateway lets you set per-tool permissions (e.g., an agent can read customer records but not modify billing) and get a log of every call, which matters the moment a customer asks what data an AI touched or a security review asks the same question. For a 10-200 person company without a dedicated security team, this shifts agent governance from a custom-built afterthought to a configuration you turn on, provided you're already on AWS or willing to route agent traffic through Bedrock.
AWS Shows How to Cut RAG Token Costs on Bedrock by Trimming Irrelevant Context
If your support bot, sales assistant, or internal knowledge search runs on a retrieval-augmented pipeline through Bedrock (or a similar architecture), the token bill scales with how much irrelevant context gets stuffed into every prompt — long documents, boilerplate, and near-duplicate passages you retrieve 'just in case.' Query-aware compression addresses that by filtering retrieved chunks against the actual question before they reach the model, which is the same lever that determines whether a 20-person support team's AI assistant costs $200 or $2,000 a month at scale. Teams already running RAG in production should treat this as a concrete cost-reduction checklist item, not a future upgrade — it requires no model swap, only a compression step inserted into the existing retrieval-to-generation pipeline.
AWS Adds Cross-Region Routing for GPT-5.6 on Bedrock
If your support bot, lead-qualification agent, or ops automation calls GPT-5.6 through Amazon Bedrock, this removes a real operational headache: capacity crunches in a single region that cause dropped or delayed responses during peak hours. Instead of writing and maintaining your own retry-and-failover logic across regions, Bedrock now handles that routing for you, which means fewer 3am pages when a customer-facing AI workflow starts throttling. Teams running lean ops (10-200 people) rarely have spare engineering time to build resilience infrastructure themselves, so this is a case where the cloud provider absorbing that complexity is a direct, if modest, win for uptime of any AI-driven sales or support pipeline built on Bedrock.
AWS Lets Agent Builders Restrict Web Search to Approved, Recent Sources
For a 10-200 person B2B company running a support or sales agent that pulls live web results to answer customer questions, this closes a real gap: until now, an agent grounded in open web search could just as easily surface a three-year-old blog post or a competitor's page as your own documentation. Teams building on AgentCore can now lock search to a whitelist (docs.yourcompany.com, trusted partner sites, industry standards bodies) and require content published within a set window, which matters for anything involving pricing, compliance, or product specs that change often. It also gives ops and legal teams a concrete control to point to when a customer or auditor asks how the agent decided what to cite, rather than an unverifiable 'it searched the web.'
Free AI Diagnostic
Fifteen minutes, no email required. It maps where your work actually goes and ranks what is worth automating first.
Start the free diagnosticStarts immediately in the browser.
- Fee
- Free
- Length
- 15 minutes
You keep the ranked list of candidates either way.