<?xml version="1.0" encoding="utf-8"?>
<feed xmlns="http://www.w3.org/2005/Atom">
  <title>Off-by-none Serverless Newsletter</title>
  <subtitle>Stay up to date on using serverless to build modern applications in the cloud. Get insights from experts, product releases, industry happenings, tutorials and much more, every week!</subtitle>
  <link href="https://offbynone.io/feed/" rel="self"/>
  <link href="https://offbynone.io/"/>
  <updated>2026-07-30T12:14:56Z</updated>
  <id>https://offbynone.io/</id>
  <author>
    <name>Jeremy Daly</name>
    <email>contact@jeremydaly.com</email>
  </author>
  <entry>
    <title>Issue #371: Could your SaaS be replaced by a Markdown File? 📝</title>
    <link href="https://offbynone.io/issues/371/"/>
    <updated>2026-07-07T12:00:00Z</updated>
    <summary>In this issue, Claude Cowork breaks free of your laptop, MiniMax lands on Bedrock, and we weigh what taste and judgment are still worth.</summary>
    <id>https://offbynone.io/issues/371/</id>
    <content type="html">&lt;h2&gt;Could your SaaS be replaced by a Markdown File? 📝&lt;/h2&gt;
&lt;p&gt;In our &lt;a href=&quot;https://offbynone.io/issues/370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;previous issue&lt;/a&gt;, Anthropic launched Claude Sonnet 5, CloudFormation got much faster, and OpenAI started making Jalapeños. In this issue, Claude Cowork breaks free of your laptop, MiniMax lands on Bedrock, and we weigh what taste and judgment are still worth. Plus, we have lots of awesome content from the cloud, serverless, and AI communities.&lt;/p&gt;
&lt;h3&gt;News &amp;amp; Announcements&lt;/h3&gt;
&lt;p&gt;Most of the interesting news from this week is around agent plumbing, and AWS still seems to be doing a lot of that work. The most interesting is &lt;a href=&quot;https://aws.amazon.com/blogs/machine-learning/structured-memory-filtering-with-metadata-in-agentcore-memory?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23371&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;structured memory filtering with metadata in AgentCore Memory&lt;/a&gt;, which layers attribute-based filtering on top of namespace isolation. Memory scope and attributes are two different things, and it&#39;s nice to see that AWS is getting this one right. AgentCore also &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/07/amazon-bedrock-agentcore-increases-default-runtime-quota-limits?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23371&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;bumped its default runtime quotas&lt;/a&gt;, now up to 200 agent interactions and 25 new sessions per second, with US East and West carrying 5,000 concurrent sessions. More headroom is good, just be sure to use it wisely.&lt;/p&gt;
&lt;p&gt;On the model side, &lt;a href=&quot;https://aws.amazon.com/blogs/machine-learning/run-minimax-models-on-amazon-bedrock?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23371&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;MiniMax models are now on Amazon Bedrock&lt;/a&gt; across 14 regions. I like the MiniMax models, and with tool-calling, implicit prompt caching, and $0.30 price tag per million input tokens, it could be a nice drop-in for your agentic workflows. AWS also shipped an &lt;a href=&quot;https://aws.amazon.com/blogs/opensource/introducing-mcp-server-for-registry-of-open-data-on-aws?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23371&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;open source MCP server for the Registry of Open Data&lt;/a&gt;, giving AI assistants access to more than 1,100 public datasets for natural-language discovery. Useful for anyone who wants to ground their research against satellite imagery, climate, genomics data, etc. without knowing the catalog cold.&lt;/p&gt;
&lt;p&gt;For the folks who actually keep things running in the cloud, ECS picked up &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/07/amazon-ecs-aws-management-console?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23371&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;real-time deployment observability in the console&lt;/a&gt; with a live timeline, circuit breaker monitoring, and failed-task diagnostics at no extra charge, which is exactly the kind of thing you appreciate at 2 a.m. while watching a rollout stall. And &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/07/cognito-provisioned-limits?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23371&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Cognito moved API rate limits to self-service&lt;/a&gt;, so you can raise them from the console up to your account maximum instead of opening a Service Quotas ticket and waiting.&lt;/p&gt;
&lt;p&gt;Outside of AWS, &lt;a href=&quot;https://claude.com/blog/cowork-web-mobile?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23371&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Claude Cowork landed on web and mobile&lt;/a&gt; with background execution and scheduled runs, so a task can prep your morning briefing at 6am and ping your phone when it needs a decision before continuing. That was a big theme at the AI Engineer World&#39;s Fair last week, workflows moving off your laptop so you&#39;re no longer limited by compute (just money). &lt;a href=&quot;https://siliconangle.com/2026/07/01/together-ai-raises-800m-grow-ai-optimized-public-cloud?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23371&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Together AI raised $800M&lt;/a&gt; at an $8.3B valuation to grow its AI-optimized public cloud. They claim their ATLAS technology speeds up some inference workloads by as much as 400%.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://vercel.com/blog/dockerfile-on-vercel?ref=console.dev&amp;utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23371&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Vercel will now run any Dockerfile&lt;/a&gt; on its Fluid compute platform, so Go, Rails, or Spring Boot deploy with the same preview-and-scale workflow as everything else. And &lt;a href=&quot;https://eventcatalog.dev/blog/eventcatalog-v4?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23371&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;EventCatalog v4&lt;/a&gt; reframes itself from event docs to a broader architecture catalog, adding a Systems resource type and an agent that keeps your docs synced to the codebase.&lt;/p&gt;
&lt;h3&gt;Tutorials&lt;/h3&gt;
&lt;ul&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/machine-learning/teaching-models-to-forget-selective-unlearning-with-amazon-nova?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23371&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Teaching models to forget: Selective unlearning with Amazon Nova&lt;/a&gt; by Qian Hu&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/machine-learning/build-a-serverless-image-editing-agent-with-amazon-bedrock-agentcore-harness?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23371&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Build a serverless image editing agent with Amazon Bedrock AgentCore harness&lt;/a&gt; by Salman Ahmed&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/machine-learning/automatically-redact-pii-in-images-with-amazon-nova?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23371&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Automatically redact PII in images with Amazon Nova&lt;/a&gt; by Caroline Des Rochers&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/database/data-masking-in-amazon-rds-for-oracle?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23371&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Data masking in Amazon RDS for Oracle&lt;/a&gt; by Jobin Joseph&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://claude.com/blog/claude-model-and-effort-level-in-claude-code?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23371&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Choosing a Claude model and effort level in Claude Code&lt;/a&gt; by Lydia Hallie&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://dev.to/aws/i-built-pr-preview-environments-with-aws-lambda-microvms-and-cut-staging-costs-by-78-2d3i?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23371&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;I Built PR Preview Environments With AWS Lambda MicroVMs and Cut Staging Costs by 78%&lt;/a&gt; by Jatin Mehrotra&lt;/li&gt;
&lt;/ul&gt;
&lt;h3&gt;Reads&lt;/h3&gt;
&lt;p&gt;&lt;a href=&quot;https://www.linkedin.com/pulse/ten-years-micro-frontend-decisions-condensed-skill-luca-mezzalira-lfqge?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23371&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Ten years of micro-frontend decisions, condensed into a skill&lt;/a&gt; by Luca Mezzalira&lt;br /&gt;
Luca took a decade of hard-won micro-frontend lessons and packed them into a skill that keeps AI agents from wrecking your boundaries. If you&#39;ve ever watched a cross-boundary import or a shared global sneak into a codebase, this is the guardrail you wish you&#39;d had.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://www.swyx.io/america250?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23371&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;What America has meant to me&lt;/a&gt; by Shawn &amp;quot;swyx&amp;quot; Wang&lt;br /&gt;
Swyx walks through his three &amp;quot;runs&amp;quot; in America across almost twenty years, from finance to engineering to founding Latent Space and the AI Engineer movement.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://caylent.com/blog/claude-sonnet-5-launch-analysis-what-changed-what-matters-and-what-to-validate?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23371&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Claude Sonnet 5 Launch Analysis: What Changed, What Matters, and What to Validate&lt;/a&gt; by Guille Ojeda&lt;br /&gt;
Guille breaks down what actually changed in Sonnet 5, including the adaptive thinking defaults, effort controls, and a tokenizer shift that bumps your token counts by about 30%. The real takeaway is to validate against your own workloads instead of trusting the benchmarks.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://charity.wtf/p/in-defense-of-ai-mandates?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23371&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;In defense of AI mandates&lt;/a&gt; by Charity Majors&lt;br /&gt;
Charity argues that top-down AI mandates give managers the cover they need to actually spend time and budget on adoption. Her point is that companies have to decide whether AI is existential or optional, then fund that decision like they mean it.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://claude.com/blog/a-field-guide-to-claude-fable-finding-your-unknowns?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23371&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;A Field Guide to Claude Fable: Finding Your Unknowns&lt;/a&gt; by Thariq Shihipar&lt;br /&gt;
Thariq lays out a framework for categorizing what you don&#39;t know and applying the right pattern at each stage of an AI coding workflow. The blind spot passes and pre-implementation planning are the parts you should steal first.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://lucvandonkersgoed.com/2026/07/01/i-replaced-my-github-runners-with-lambda-microvms-and-maybe-you-should-too?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23371&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;I replaced my GitHub runners with Lambda MicroVMs, and maybe you should too&lt;/a&gt; by Luc van Donkersgoed&lt;br /&gt;
Luc swaps GitHub-hosted runners for Lambda MicroVMs and is honest about the tradeoffs, which land at minimal cost savings but a few minutes faster per run. The real cost is the operational overhead of owning your runner infrastructure, and he doesn&#39;t pretend otherwise.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://dev.to/aws-builders/what-if-building-an-ai-chatbot-was-as-easy-as-snapping-lego-bricks-together-aws-blocks-3lhm?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23371&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;What if building an AI chatbot was as easy as snapping LEGO bricks together? - AWS Blocks&lt;/a&gt; by Luis Fernando de León Ramírez&lt;br /&gt;
Luis shows off AWS Blocks, the new AWS open-source framework that turns TypeScript into AWS serverless services with a snap-together feel. The walkthrough covers local dev, sandbox testing, IAM, and shipping a Bedrock Nova Lite chatbot with real code.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://dev.to/aws-builders/lambda-microvms-vs-agentcore-runtime-when-to-use-each-for-production-agents-5gm7?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23371&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Lambda MicroVMs vs AgentCore Runtime: When to Use Each for Production Agents&lt;/a&gt; by Gerardo Arroyo&lt;br /&gt;
Gerardo lays out when to reach for Lambda MicroVMs versus AgentCore Runtime for production agents, and where the two actually complement each other. The framing around coding agents that need secure execution sandboxes is the useful bit.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://medium.com/@DaveThackeray/15-things-i-learned-at-ai-engineer-worlds-fair-2026-48b66456ebc1?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23371&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;15 things I learned at AI Engineer World’s Fair 2026&lt;/a&gt; by Dave Thackeray&lt;br /&gt;
Dave&#39;s fifteen takeaways cover the mismatch between probabilistic models and deterministic infra, context cost, and patterns like semantic routing and post-generation veto systems. Since I was there too, I can tell you his list holds up.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://tylerfolkman.substack.com/p/the-ai-chatbot-era-is-ending-teams?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23371&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;The AI Chatbot Era Is Ending. Teams Are Optimizing the Wrong Layer.&lt;/a&gt; by Tyler Folkman&lt;br /&gt;
Tyler Folkman argues that teams should stop optimizing prompts and start designing delegation frameworks for AI agents. He presents a five-layer Delegation Stack based on hundreds of agent sessions, backed by recent data from OpenAI and Anthropic showing enterprise shifts toward agent-based workflows.&lt;/p&gt;
&lt;h3&gt;Podcasts, Videos, and more&lt;/h3&gt;
&lt;p&gt;&lt;a href=&quot;https://www.lennysnewsletter.com/p/sonnet-5-review-i-ran-64-generations?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23371&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Sonnet 5 review: I ran 64 generations to find out if it&#39;s worth it&lt;/a&gt;&lt;br /&gt;
Claire ran 64 generations pitting Sonnet 5 against Sonnet 4.6, Opus 4.8, GPT-5.5, and Gemini 3 Pro with her own eval framework. She scored PRD quality, prototype generation, agentic completion, and agent personality, and the results surprised her.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://www.youtube.com/@aiDotEngineer/streams?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23371&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AI Engineer - YouTube&lt;/a&gt;&lt;br /&gt;
The AI Engineer channel is where the World&#39;s Fair sessions land once they&#39;re posted, so subscribe if you couldn&#39;t make it in person. Right now, all the livestreams of the keynotes and main stage track are posted, so check those out when you get some time.&lt;/p&gt;
&lt;h3&gt;New from AWS&lt;/h3&gt;
&lt;ul&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/07/impact-analysis-aws-security-hub?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23371&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS Security Hub adds impact analysis for exposure findings&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/07/secrets-manager-managed-external-secrets-paddle-gitlab?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23371&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS Secrets Manager adds managed external secrets support for Paddle and GitLab&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/07/amazon-cloudwatch-log-alarms?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23371&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon CloudWatch supports creating alarms from log queries&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/cloudwatch-dynamic-instrumentation?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23371&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;CloudWatch Application Signals now supports Dynamic Instrumentation&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/aws-config-new-resource-types?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23371&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS Config now supports 8 new resource types&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/07/amazon-ecs-express-mode-custom-task-def?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23371&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon ECS Express Mode now supports custom task definitions&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/07/amazon-opensearch-service-optimized-log-analytics?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23371&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon OpenSearch Service optimized for log analytics&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/07/amazon-ecs-circuit-breaker-settings?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23371&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon ECS now supports configurable deployment circuit breaker settings&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/07/amazon-emr-serverless?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23371&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon EMR Serverless now supports larger worker sizes to run more compute and memory intensive workloads&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/07/amazon-eks-auto-mode-gpu-price?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23371&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon EKS Auto Mode reduces GPU management fees by up to 60%&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/07/amazon-ecs-managed-instances-gpu-price?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23371&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon ECS Managed Instances reduces GPU management fees by up to 60%&lt;/a&gt;&lt;/li&gt;
&lt;/ul&gt;
&lt;h3&gt;Thoughts from Social&lt;/h3&gt;
&lt;blockquote class=&quot;twitter-tweet&quot;&gt;&lt;p lang=&quot;en&quot; dir=&quot;ltr&quot;&gt;“Half of the companies here at &lt;a href=&quot;https://x.com/aiDotEngineer?ref_src=twsrc%5Etfw&quot;&gt;@aiDotEngineer&lt;/a&gt; could be a markdown file!” ~ &lt;a href=&quot;https://x.com/theo?ref_src=twsrc%5Etfw&quot;&gt;@theo&lt;/a&gt; 🌶️🌶️🌶️ &lt;a href=&quot;https://t.co/kxaTCZUd2G&quot;&gt;pic.twitter.com/kxaTCZUd2G&lt;/a&gt;&lt;/p&gt;&amp;mdash; Jeremy Daly (@jeremy_daly) &lt;a href=&quot;https://x.com/jeremy_daly/status/2072830651846529317?ref_src=twsrc%5Etfw&quot;&gt;July 2, 2026&lt;/a&gt;&lt;/blockquote&gt; &lt;script async=&quot;&quot; src=&quot;https://platform.x.com/widgets.js&quot; charset=&quot;utf-8&quot;&gt;&lt;/script&gt;
&lt;h3&gt;Developer Tools&lt;/h3&gt;
&lt;p&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/machine-learning/simplify-model-selection-in-amazon-bedrock-with-the-open-source-model-profiler?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23371&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Simplify model selection in Amazon Bedrock with the open source Model Profiler&lt;/a&gt; by Maria Oliva Calero&lt;br /&gt;
Maria walks through the Bedrock Model Profiler, an open-source tool that pulls 120+ foundation models into one searchable view. The serverless pipeline stitches together five AWS APIs and two public sources so you can filter on pricing, region, quotas, and lifecycle without tab-hopping.&lt;/p&gt;
&lt;h3&gt;Final Thoughts 🤔&lt;/h3&gt;
&lt;p&gt;Theo&#39;s comment that half the companies at the AI Engineer World&#39;s Fair could be a markdown file landed hard for some. Because he&#39;s not entirely wrong. Point a capable model at a well-structured markdown file, wrap it in a skill and hand it to Claude, and a lot of what those startups demoed just falls out the other end. Instruction-following has gotten good enough that the markdown file basically becomes the product.&lt;/p&gt;
&lt;p&gt;The catch is that a markdown file captures what you want done, not the judgment behind how it&#39;s done. Romain Huet from OpenAI made this point in his keynote: engineering has always been about solving problems by combining the latest science &amp;quot;with design, with taste, with judgment, and most of all, imagination&amp;quot; to make something people can actually use. That judgment is what a lot of SaaS companies are actually selling. They&#39;ve spent years sequencing business processes and getting them battle-tested across thousands of customers, and the good startups are staffed by domain experts who already worked out the right way to handle a problem so you don&#39;t have to. Replace that with a markdown file and you&#39;re trading years of accumulated judgment for something you wrote in an afternoon.&lt;/p&gt;
&lt;p&gt;Then there&#39;s the bill. Running a skill once is cheap and kind of magical. Running it thousands of times a day against a frontier model is a line item nobody budgeted for. The workflows that actually scale are mostly deterministic, with the model reserved for the handful of points where a real decision has to be made: read this data, decide whether to proceed, escalate to a human, or route somewhere else. That&#39;s where an LLM earns its cost. Wrapping the whole process in one because you can is how you light money on fire.&lt;/p&gt;
&lt;p&gt;And almost all of it shows up in the last 10%. AI might get you 90% of the way to a working product in a weekend, and that first 90% is impressive. But the last 10% is the edge cases, the industry quirks, and the hard-won defaults that come from domain experts and a team of PMs and engineers talking to real customers. You&#39;ll build the features you use every day and feel great about it, right up until you hit the edge cases and nuances that really matter. That gap is brutal, and no markdown file is going to close it for you.&lt;/p&gt;
&lt;p&gt;So could your SaaS be replaced by a markdown file? For a thin wrapper, sure, and probably very soon. For anything that encodes real judgment about a hard problem, the markdown file gets you a convincing demo and a bill for the 10% you didn&#39;t build. Know which one you&#39;re building.&lt;/p&gt;
&lt;p&gt;See you next week,&lt;br /&gt;
Jeremy&lt;/p&gt;
&lt;hr /&gt;
&lt;p&gt;I hope you enjoyed this newsletter. We&#39;re always looking for ideas and feedback to make it better and more inclusive, so please feel free to reach out to me via &lt;a href=&quot;https://bsky.app/profile/jeremydaly.com&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Bluesky&lt;/a&gt;, &lt;a href=&quot;http://www.linkedin.com/in/jeremydaly&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;LinkedIn&lt;/a&gt;, &lt;a href=&quot;http://twitter.com/jeremy_daly&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;X&lt;/a&gt;, or &lt;a href=&quot;mailto:contact@jeremydaly.com&quot;&gt;email&lt;/a&gt;.&lt;/p&gt;
</content>
  </entry>
  <entry>
    <title>Issue #370: Live from AI Engineer World&#39;s Fair 🎡</title>
    <link href="https://offbynone.io/issues/370/"/>
    <updated>2026-06-30T12:00:00Z</updated>
    <summary>In this issue, Anthropic launches Claude Sonnet 5, CloudFormation gets much faster, and OpenAI starts making Jalapeños.</summary>
    <id>https://offbynone.io/issues/370/</id>
    <content type="html">&lt;h2&gt;Live from AI Engineer World&#39;s Fair 🎡&lt;/h2&gt;
&lt;p&gt;In our &lt;a href=&quot;https://offbynone.io/issues/369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;previous issue&lt;/a&gt;, AWS Summit NYC unsurprisingly went heavy on AI agents, Lambda picked up MicroVMs for isolated sandboxes, and AWS Blocks brought IfC back into the conversation. This week, Anthropic launches Claude Sonnet 5, CloudFormation gets much faster, and OpenAI starts making Jalapeños. Plus, we have plenty of excellent cloud, serverless, and AI content from the community.&lt;/p&gt;
&lt;h3&gt;News &amp;amp; Announcements&lt;/h3&gt;
&lt;p&gt;I&#39;m hanging out at the AI Engineer World&#39;s Fair this week in San Francisco, and the vibe here is amazing. We&#39;ll get to that in a minute. The big headline is Claude Sonnet 5. Anthropic &lt;a href=&quot;https://www.anthropic.com/news/claude-sonnet-5?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;shipped it&lt;/a&gt; with a real jump in coding and agentic performance at Sonnet pricing, with introductory rates of $2 and $10 per million tokens through August 2026. It&#39;s &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/claude-sonnet-5-now-available-on-aws?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;already on AWS&lt;/a&gt; via Bedrock and the Claude Platform, and Aamna Najmi&#39;s &lt;a href=&quot;https://aws.amazon.com/blogs/machine-learning/introducing-claude-sonnet-5-on-aws-anthropics-most-capable-sonnet-model?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;walkthrough&lt;/a&gt; has the SDK and Converse examples.&lt;/p&gt;
&lt;p&gt;Anthropic paired that with platform news: &lt;a href=&quot;https://claude.com/blog/claude-in-microsoft-foundry?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Claude is GA in Microsoft Foundry&lt;/a&gt; (Opus 4.8 and Haiku 4.5, Azure- or Anthropic-hosted), a self-hosted &lt;a href=&quot;https://claude.com/blog/introducing-the-claude-apps-gateway?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Claude apps gateway&lt;/a&gt; that adds SSO and centralized policy to Claude Code on Bedrock and GCP, and &lt;a href=&quot;https://claude.com/blog/agent-identity-access-model?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;agent identity&lt;/a&gt;, which finally gives Claude Tag its own workspace account instead of borrowing a human&#39;s.&lt;/p&gt;
&lt;p&gt;On AWS, CloudFormation got faster with &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/aws-cloudformation-cdk?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Express mode&lt;/a&gt; speeding up stack operations up to 4x by returning once a configuration is applied (&lt;a href=&quot;https://aws.amazon.com/blogs/aws/accelerate-your-infrastructure-deployments-by-up-to-4x-with-aws-cloudformation-express-mode?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Channy&#39;s writeup&lt;/a&gt; covers the tradeoffs), and &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/aws-cloudformation?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;pre-deployment validation&lt;/a&gt; now catches quota, Config, and ECR issues before provisioning. Faster loops and fewer failed deploys is what the codegen era needs out of IaC. &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/amazon-elasticache-valkey-9-1?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;ElastiCache added Valkey 9.1&lt;/a&gt; with a new I/O threading model and commands like HGETDEL (more details &lt;a href=&quot;https://aws.amazon.com/blogs/database/announcing-valkey-9-1-for-amazon-elasticache?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;here&lt;/a&gt;).&lt;/p&gt;
&lt;p&gt;Also note that AWS is &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/aws-service-availability?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;moving several services to maintenance mode&lt;/a&gt; on July 30, including Bedrock Agents (now Bedrock Agents Classic), Kendra, and Q Business. Existing customers can stay, but if you just haven&#39;t gotten around to trying out myApplications on the AWS Console yet, you&#39;re out of luck. 😉 Glad to see that AWS continues to trim some fat, but as I&#39;ve said before, they should have taken the whole leg at once instead of just cutting out the rot. That&#39;s a horrible visual, but an apt analogy IMO.&lt;/p&gt;
&lt;p&gt;Outside AWS, Oracle opened up &lt;a href=&quot;https://blogs.oracle.com/mysql/the-next-phase-of-mysql-community-engagement-accelerating-participation-and-collaboration?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;MySQL governance&lt;/a&gt; with AWS and Google Cloud on the steering committee, OpenAI and Broadcom &lt;a href=&quot;https://openai.com/index/openai-broadcom-jalapeno-inference-chip?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;unveiled Jalapeño&lt;/a&gt;, a from-scratch LLM inference chip already running GPT-5.3-Codex-Spark, and OpenAI &lt;a href=&quot;https://openai.com/index/previewing-gpt-5-6-sol?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;previewed GPT-5.6&lt;/a&gt; in three variants with a phased rollout. I know some people with access and now I&#39;m super jealous.&lt;/p&gt;
&lt;h3&gt;Tutorials&lt;/h3&gt;
&lt;ul&gt;
&lt;li&gt;&lt;a href=&quot;https://claude.com/blog/getting-started-with-loops?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Getting started with loops&lt;/a&gt; by Anthropic&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/machine-learning/build-generative-ui-for-ai-agents-on-amazon-bedrock-agentcore-with-the-ag-ui-protocol?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Build generative UI for AI agents on Amazon Bedrock AgentCore with the AG-UI protocol&lt;/a&gt; by Ryan Razkenari&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/machine-learning/fine-tune-amazon-nova-models-for-accurate-email-data-extraction?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Fine-tune Amazon Nova models for accurate email data extraction&lt;/a&gt; by Le Vy&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/machine-learning/pair-nova-2-lite-with-claude-for-cost-optimized-document-processing?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Pair Nova 2 Lite with Claude for cost-optimized document processing&lt;/a&gt; by Sanghwa Na&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/machine-learning/how-loka-built-a-natural-low-latency-voice-agent-with-amazon-nova-2-sonic?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;How Loka Built a Natural, Low-Latency Voice Agent with Amazon Nova 2 Sonic&lt;/a&gt; by Bojan Jakimovski&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/database/user-authentication-and-session-management-with-amazon-aurora-dsql?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;User authentication and session management with Amazon Aurora DSQL&lt;/a&gt; by Chaitanya chary Chatlapally&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://dev.to/mierune/aws-amplify-lambda-microvms-a-serverless-linux-desktop-1eif?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS Amplify + Lambda MicroVMs = A Serverless Linux Desktop!&lt;/a&gt; by Kanahiro Iguchi&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://dev.to/aws-builders/i-made-an-aws-lambda-microvm-publicly-accessible-for-0month-heres-the-full-setup-36fn?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;I made an AWS Lambda MicroVM publicly accessible for $0/month (here&#39;s the full setup)&lt;/a&gt; by Alexey Vidanov&lt;/li&gt;
&lt;/ul&gt;
&lt;h3&gt;Reads&lt;/h3&gt;
&lt;p&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/architecture/lessons-learned-from-scaling-to-1-million-lambda-functions?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Lessons learned from scaling to 1 million Lambda functions&lt;/a&gt; by Ben Freiberg&lt;br /&gt;
A million Lambda functions across thousands of accounts is the kind of scale that breaks tooling nobody expects to break, and CloudFormation StackSets is right at the top of that list. Read it for the observability cost lessons, which is where most teams get caught off guard.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://www.readysetcloud.io/blog/allen.helton/that-probably-doesnt-need-to-be-saas?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;PSA: That probably doesn&#39;t need to be SaaS | Ready, Set, Cloud!&lt;/a&gt; by&lt;br /&gt;
Allen&#39;s point about builders shipping products now instead of writing up what they learned is one I&#39;ve felt myself lately. The tradeoff he names, easier building at the cost of shared knowledge, is real and has stuck with me.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://charity.wtf/2026/06/24/make-ai-boring-again-xpost?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Make AI Boring Again&lt;/a&gt; by Charity Majors&lt;br /&gt;
Charity&#39;s case for learning AI so you understand how it fails, rather than opting out, is the right instinct for engineers. She doesn&#39;t wave away the real problems around training data, labor, and energy, which is what makes the argument land.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://fortem.dev/blog/fargate-vs-lambda?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS Fargate vs Lambda: When Does Lambda Stop Being Cheaper?&lt;/a&gt; by Matt S&lt;br /&gt;
The useful reframe here is that Lambda&#39;s breakeven is driven by execution duration rather than request volume, which is backwards from how most people reason about it. A 200ms API staying cheaper up to 6-8M calls a month is a handy number to keep in your back pocket.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://studyfromexperts.com/blogs/why-we-built-our-own-crm-for-under-5-using-aws-kiro?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Why We Built Our Own CRM for Under $5 using AWS Kiro&lt;/a&gt; by Lee Gilmore&lt;br /&gt;
Lee provides a solid writeup advocating build versus buy, with a clever DynamoDB-to-Aurora DSQL sync using change data capture. Just remember the few dollars a month doesn&#39;t include the time you&#39;ll spend maintaining it, which is the part that&#39;ll bite you. But I&#39;d probably build this myself too. 😂&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://javierv.dev/blog/why-i-still-approve-my-memory-by-hand/?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Why I still approve my memory by hand&lt;/a&gt; by Javier Villanueva&lt;br /&gt;
The argument that human-in-the-loop curation beats automated validation for a single-user knowledge base is the pattern I recommend. Automated approval mostly gives you correlated bias dressed up as confirmation, which is a trap you don&#39;t want at this scale.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://www.allthingsdistributed.com/2026/06/return-to-two-pizza-culture.html?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;A return to two-pizza culture&lt;/a&gt; by Dr. Werner Vogels&lt;br /&gt;
Werner tying two-pizza teams to AI agents is a sharp framing, and the Quick Desktop story of an overnight prototype reshaping how the team planned is the example that sells it. The part I&#39;d watch is what happens to documentation once prototypes get this cheap.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://newsletter.pragmaticengineer.com/p/impressions-from-visiting-openai?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Impressions from visiting OpenAI, Anthropic, &amp;amp; Cursor&lt;/a&gt; by Gergely Orosz&lt;br /&gt;
Gergely&#39;s four trends are a good pulse check, especially cloud agents going mainstream and engineers optimizing their code for agent efficiency. The cost-reduction pressure he describes is what I&#39;d keep an eye on, since it shapes what these tools become next.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://theburningmonk.com/2026/06/what-you-need-to-know-about-lambda-microvms?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;What you need to know about Lambda MicroVMs&lt;/a&gt; by Yan Cui&lt;br /&gt;
Yan&#39;s framing is the clearest I&#39;ve seen: MicroVMs sit closer to EC2 than Lambda, since you&#39;re running persistent VMs and managing the fleet yourself. If you came in expecting request/response Lambda ergonomics, read this first to reset your expectations.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://dev.to/coa00/i-tried-aws-blocks-on-a-real-amplify-gen2-project-local-dynamodb-no-aws-account-1-second-loops-4gm3?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;I Tried AWS Blocks on a Real Amplify Gen2 Project — Local DynamoDB, No AWS Account, 1-Second Loops&lt;/a&gt; by Kohei Aoki&lt;br /&gt;
A hands-on look at AWS Blocks with simulated local DynamoDB and one-second feedback loops instead of cloud deploy cycles. The fast feedback is a nice side-benefit, but the real win is not having to bifurcate your business logic into IaC.&lt;/p&gt;
&lt;h3&gt;Podcasts, Videos, and more&lt;/h3&gt;
&lt;p&gt;&lt;a href=&quot;https://www.youtube.com/watch?v=paoEOWbyBxE?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Introducing AWS Lambda MicroVMs | Serverless Office Hours&lt;/a&gt;&lt;br /&gt;
The Serverless Office Hours crew demos MicroVMs live, which is the fastest way to see snapshot launches and suspend/resume in action. If you want the mechanics behind the sandbox-per-session pattern, start here before the docs.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://www.lennysnewsletter.com/p/glm-52-why-im-replacing-opus-in-claude?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;GLM 5.2: why I’m replacing Opus in Claude Code with this new model&lt;/a&gt;&lt;br /&gt;
Claire&#39;s walkthrough of dropping GLM 5.2 into Claude Code is a useful look at what open-weight actually buys you on cost and vendor independence. She&#39;s clear about where it fell short too, which keeps it from being a hype piece.&lt;/p&gt;
&lt;h3&gt;New from AWS&lt;/h3&gt;
&lt;ul&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/aws-waf-amazon-bedrock-agentcore?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS WAF adds support for Amazon Bedrock AgentCore Gateway&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/amazon-s3-cloudwatch-logs-tables?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon S3 server access logs now deliver to Amazon CloudWatch Logs and Amazon S3 Tables&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/aws-backup-amazon-s3-copy-enhancement?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS Backup enhances Amazon S3 backup copy performance&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/amazon-cloudwatch-logs-resource-tags?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon CloudWatch Logs enriches log events with AWS resource tags&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/amazon-cognito-customer-managed-key?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon Cognito now supports customer managed key for encryption at rest&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/amazon-workspaces-ai?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Announcing general availability of Amazon WorkSpaces for AI agents&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/amazon-sagemaker-ai-gemma-4?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon SageMaker AI now supports serverless model customization for Gemma 4 models&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/amazon-mwaa-serverless-vpc?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon MWAA Serverless now supports shared VPC configurations&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/amazon-route-53-global-resolver?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon Route 53 Global Resolver now supports sharing DNS Views between AWS Accounts&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/amazon-emr-serverless?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon EMR Serverless now supports live configuration updates without application restarts&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/claude-tag-aws-marketplace?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Claude Tag is now available in beta via Claude Enterprise in AWS Marketplace&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/amazon-opensearch-service-ai-migrations?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23370&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon OpenSearch Service now offers AI-assisted migrations&lt;/a&gt;&lt;/li&gt;
&lt;/ul&gt;
&lt;h3&gt;Final Thoughts 🤔&lt;/h3&gt;
&lt;p&gt;If 2025 was the year of agents, 2026 is the year of loops and software factories. That was the throughline at the AI Engineer World&#39;s Fair today, where more than 7,000 engineers gathered in San Francisco to compare notes and trade war stories. Shawn &amp;quot;swyx&amp;quot; Wang set the stage with a talk about Loopcraft, tracing how the loops keep compounding until you reach the highest one: engineers learning from each other. Peter Steinberger, the creator of Open Claw, and probably several steps ahead of most, put the operational edge on it: keeping ten terminals open to babysit your agents is already the old way. An agent manager that lets you drop into session when you need to take control is what comes next.&lt;/p&gt;
&lt;div style=&quot;border:solid 1px #cccccc;margin-bottom:30px;&quot;&gt;&lt;img alt=&quot;AI Engineer World&#39;s Fair&quot; src=&quot;https://offbynone.io/images/issues/ai-engineer-worlds-fair-2026.jpg&quot; class=&quot;img-fluid&quot; width=&quot;100%&quot; /&gt;&lt;/div&gt;
&lt;p&gt;The energy around that vision is hard to miss, and the pace backs it up. Romain Huet and Alexander Embiricos from OpenAI mentioned they&#39;re shipping new models roughly every six weeks now, a cadence that would have sounded absurd a year ago. And it isn&#39;t only the frontier labs. Zixuan Li from &lt;a href=&quot;http://z.ai/&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Z.ai&lt;/a&gt; dialed in to show off GLM 5.2, an open-weight model built for long-horizon tasks that&#39;s pretty close to Opus 4.8 and GPT 5.5 on the benchmarks. Claire Vo&#39;s video on swapping Opus for GLM 5.2 in Claude Code says it holds up in real work too, and it has me tempted to throw an RTX 5090 in an Ubuntu box and run my own local AI lab. Capability is getting faster, cheaper, and more portable by the month.&lt;/p&gt;
&lt;p&gt;But the cracks are starting to show, and it&#39;s in the same place it always is: software maintenance. The software-factory pitch was great for greenfield, and one-shotting a brand new app is exactly where these current models shine. Dexter Horthy hammered on this in his harness talk, that maintaining all the AI slop we&#39;re generating starts to break down after only a few months, and there&#39;s still a stubborn list of problems the agents can&#39;t solve without human intervention. The models are getting great at producing something from nothing, yet still struggle the moment you point them at a large, living codebase, including the ones it generated from scratch.&lt;/p&gt;
&lt;p&gt;That&#39;s the gap I keep coming back to. Faster models, open weights, even a local lab of my own, none of it touches the part where the code you wrote three months ago now needs maintenance, new features, and security/performance upgrades. Loops are fantastic at the start of a project and shaky in the messy middle, where most useful software lives. If 2026 really is the year of software factories, the interesting work is less about building faster and more about whether the loop can survive contact with a real world software lifecycle, because that&#39;s the part nobody has solved yet.&lt;/p&gt;
&lt;p&gt;See you next week,&lt;br /&gt;
Jeremy&lt;/p&gt;
&lt;hr /&gt;
&lt;p&gt;I hope you enjoyed this newsletter. We&#39;re always looking for ideas and feedback to make it better and more inclusive, so please feel free to reach out to me via &lt;a href=&quot;https://bsky.app/profile/jeremydaly.com&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Bluesky&lt;/a&gt;, &lt;a href=&quot;http://www.linkedin.com/in/jeremydaly&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;LinkedIn&lt;/a&gt;, &lt;a href=&quot;http://twitter.com/jeremy_daly&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;X&lt;/a&gt;, or &lt;a href=&quot;mailto:contact@jeremydaly.com&quot;&gt;email&lt;/a&gt;.&lt;/p&gt;
</content>
  </entry>
  <entry>
    <title>Issue #369: Infrastructure FROM Code is Back! 🧑‍💻</title>
    <link href="https://offbynone.io/issues/369/"/>
    <updated>2026-06-23T12:00:00Z</updated>
    <summary>In this issue, AWS Summit NYC unsurprisingly goes heavy on AI agents, Lambda picks up MicroVMs for isolated sandboxes, and AWS Blocks leans into IfC.</summary>
    <id>https://offbynone.io/issues/369/</id>
    <content type="html">&lt;h2&gt;Infrastructure FROM Code is Back! 🧑‍💻&lt;/h2&gt;
&lt;p&gt;In our &lt;a href=&quot;https://offbynone.io/issues/368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;previous issue&lt;/a&gt;, the US government ordered Anthropic to pull Fable 5 and Mythos 5, AWS WAF started charging AI bots for content, and Bedrock added Grok, Gemma, and a pair of GPTs. This week, AWS Summit NYC unsurprisingly goes heavy on AI agents, Lambda picks up MicroVMs for isolated sandboxes, and AWS Blocks leans into IfC. Plus, we&#39;ve got lots of amazing cloud, serverless, and AI content from the community.&lt;/p&gt;
&lt;h3&gt;News &amp;amp; Announcements&lt;/h3&gt;
&lt;p&gt;Infrastructure FROM Code is back, baby! AWS announced &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/aws-blocks-preview?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS Blocks&lt;/a&gt;, an open-source TypeScript framework for composing application backends without wrestling with the underlying infrastructure tooling, and if the concept sounds familiar, it should. It&#39;s very close to the ideas we built into Ampt, and the thinking behind it is still as powerful as ever. You write application code, Blocks infers the services it needs, runs locally with built-in auth, and deploys to production AWS with no code changes. Seeing AWS adopt and lean into IfC development this directly is amazing to watch. It&#39;s early and still in preview, but it looks very promising and could be a really nice companion for lots of different projects.&lt;/p&gt;
&lt;p&gt;Most of what follows came out of the &lt;a href=&quot;https://aws.amazon.com/blogs/aws/top-announcements-of-the-aws-summit-in-new-york-2026?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS Summit in New York&lt;/a&gt;, and the recap is the fastest way to see the full slate in one place. The headliners are worth taking one at a time.&lt;/p&gt;
&lt;p&gt;The one I keep coming back to is &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/aws-lambda-microvms?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Lambda MicroVMs&lt;/a&gt;, a new compute primitive built on Firecracker that gives you VM-level isolation with near-instant launch and resume. You get stateful sessions that persist memory and disk for up to eight hours, full lifecycle control to launch, suspend, resume, and terminate, and support for HTTP/2, gRPC, and WebSockets. The &lt;a href=&quot;https://aws.amazon.com/blogs/aws/run-isolated-sandboxes-with-full-lifecycle-control-aws-lambda-introduces-microvms?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS blog walks through the mechanics&lt;/a&gt;: you supply a Dockerfile and a code artifact, Lambda builds a Firecracker snapshot with your app already initialized, and each user session gets its own environment with up to 16 vCPUs and 32 GB of memory. The pricing isn&#39;t great (you still pay for wall time 🤦), but this is one of the most interesting things to happen to Lambda in a while.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/amazon-bedrock-managed-knowledge-base?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon Bedrock Managed Knowledge Base is now generally available&lt;/a&gt;, a fully managed take on RAG with six native connectors (S3, SharePoint, Confluence, Google Drive, OneDrive, and a web crawler), managed vector storage, hybrid search, and an agentic retriever that can break multi-hop queries into step-by-step plans. The &lt;a href=&quot;https://aws.amazon.com/blogs/aws/introducing-amazon-bedrock-managed-knowledge-base-for-faster-more-accurate-enterprise-ai-applications?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;launch post has the details&lt;/a&gt;, including the Smart Parsing pass for content optimization. If you&#39;ve been hand-rolling RAG plumbing, this collapses a lot of it into a managed service.&lt;/p&gt;
&lt;p&gt;Bedrock AgentCore also had a big week. The &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/amazon-bedrock-agentcore-harness-generally-available?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AgentCore harness is now GA&lt;/a&gt;, a config-based path to deploying agents with managed runtime, memory strategies, multi-model support across Bedrock, OpenAI, and Gemini, and Step Functions integration, with an export path to custom code when you outgrow it (&lt;a href=&quot;https://aws.amazon.com/blogs/machine-learning/amazon-bedrock-agentcore-harness-is-now-generally-available-go-from-idea-to-production-grade-agent-in-minutes?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS frames it&lt;/a&gt; as going from idea to production agent in two API calls). &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/amazon-bedrock-agentcore-web-search?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Web Search shipped as a GA feature&lt;/a&gt; in US East (N. Virginia), giving agents real-time web access through Amazon&#39;s own index without bolting on an external provider. There are two solid reads on it from &lt;a href=&quot;https://aws.amazon.com/blogs/aws/announcing-web-search-on-amazon-bedrock-agentcore-ground-your-ai-agents-in-current-accurate-web-knowledge?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Channy&lt;/a&gt; and the &lt;a href=&quot;https://aws.amazon.com/blogs/machine-learning/introducing-web-search-on-amazon-bedrock-agentcore?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;ML blog&lt;/a&gt;, just watch out for that $7-per-1,000-queries pricing. 😬 &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/amazon-bedrock-agentcore-policy-guardrails-generally-available?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Guardrails in policy went GA&lt;/a&gt; for evaluating agent actions and blocking things like prompt injection, with policies written in natural language or code. And &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/agentcore-memory-cross-account-access?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AgentCore Memory added cross-account access&lt;/a&gt;, which sounds dull until you&#39;re trying to share memory in a multi-tenant setup. If you want a bundled view, the &lt;a href=&quot;https://aws.amazon.com/blogs/machine-learning/new-in-amazon-bedrock-agentcore-build-agents-with-broader-knowledge-and-continuous-learning?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;broader knowledge and continuous learning post&lt;/a&gt; ties the memory, web search, and paid-content pieces together.&lt;/p&gt;
&lt;p&gt;Adjacent to that, &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/amazon-bedrock-guardrails-api-ai?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Bedrock Guardrails picked up a new API aimed at agentic workflows&lt;/a&gt;. It runs in detect-only mode and returns numeric severity and confidence scores, so you set your own thresholds for blocking, retrying, or just logging at each step. Plus it hooks into agent frameworks through lifecycle hooks without making you stand up guardrail resources first. Sandeep Singh&#39;s &lt;a href=&quot;https://aws.amazon.com/blogs/machine-learning/safeguard-your-agentic-ai-applications-with-the-amazon-bedrock-guardrails-invokeguardrailchecks-api?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;walkthrough of the InvokeGuardrailChecks API&lt;/a&gt; is a useful guide if you want to dig deeper.&lt;/p&gt;
&lt;p&gt;On the storage and data side, S3 Vectors got two upgrades. It now &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/s3-vectors-supports-10000-search-results-per-query?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;returns up to 10,000 results per query&lt;/a&gt; instead of 100, with pagination so you can start processing the first page right away, and it &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/s3-vectors-reduces-query-charges-80-percent-large-indexes?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;cut query charges by up to 80% on large indexes&lt;/a&gt; (10M+ vectors), automatically across regions. S3 also &lt;a href=&quot;https://aws.amazon.com/blogs/aws/amazon-s3-annotations-attach-rich-queryable-context-directly-to-your-objects?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;added annotations&lt;/a&gt;, up to 1 GB of mutable metadata per object that surfaces automatically as queryable Iceberg tables, built for agents that need to understand data context without a human in the loop. That same theme runs through &lt;a href=&quot;https://aws.amazon.com/blogs/machine-learning/context-intelligence-for-your-data-and-ai-agents-at-scale?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS Context&lt;/a&gt;, which maps enterprise data relationships into knowledge graphs agents can query, extending the same technology already powering QuickSight.&lt;/p&gt;
&lt;p&gt;For the boring-but-useful column, &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/amazon-ecs-faster-autoscaling?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon ECS added faster service auto scaling&lt;/a&gt; with 20-second high-resolution metrics across Fargate and EC2. Channy&#39;s &lt;a href=&quot;https://aws.amazon.com/blogs/aws/amazon-ecs-introduces-new-high-resolution-metrics-for-faster-service-auto-scaling?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;breakdown&lt;/a&gt; shows scale-out dropping from over six minutes to under 90 seconds, and you can swap awkward step-scaling policies for target tracking. If you&#39;ve been overprovisioning to cover slow reactions, this is your chance to right-size.&lt;/p&gt;
&lt;p&gt;Two more preview tools from AWS push on the generated code angle. &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/aws-devops-agent-release-management?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS DevOps Agent added release management&lt;/a&gt;, which runs readiness reviews, validates infrastructure against Well-Architected practices, and generates and runs tests in isolated environments before production (the &lt;a href=&quot;https://aws.amazon.com/blogs/aws/aws-devops-agent-adds-release-management-capabilities-to-assess-code-changes-before-production-preview?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;blog has the workflow&lt;/a&gt;). And &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/aws-transform-continuous-modernization?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS Transform shipped continuous modernization&lt;/a&gt;, autonomously scanning repos to find and prioritize tech debt and opening remediation PRs, with GitHub, GitLab, and Bitbucket support (the &lt;a href=&quot;https://aws.amazon.com/blogs/aws/proactively-reduce-tech-debt-autonomously-with-aws-transform-continuous-modernization-preview?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS post&lt;/a&gt; covers end-of-life dependency detection and the Security Agent tie-in). Both are pointed squarely at the flood of AI-generated code that still needs reviewing and maintaining.&lt;/p&gt;
&lt;p&gt;Outside of AWS, Anthropic shipped a batch of updates. &lt;a href=&quot;https://claude.com/blog/claude-design-stays-on-brand-for-daily-work?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Claude Design now stays on brand&lt;/a&gt; with design-system imports from GitHub or design files, bidirectional sync with Code, and direct canvas editing. &lt;a href=&quot;https://claude.com/blog/artifacts-in-claude-code?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Claude Code picked up artifacts&lt;/a&gt; in beta for Team and Enterprise, generating shareable visual pages like incident timelines and PR walkthroughs from your codebase and conversation context. And Claude rolled out &lt;a href=&quot;https://claude.com/blog/enterprise-managed-auth?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;centrally managed authorization for MCP connectors&lt;/a&gt; using the Enterprise-Managed Authorization extension, so admins can shorten access-token lifetimes and a deprovisioned user&#39;s connector access expires fast instead of lingering. Elsewhere, &lt;a href=&quot;https://blog.cloudflare.com/temporary-accounts?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Cloudflare introduced temporary accounts for AI agents&lt;/a&gt; that let an agent deploy a Worker with &lt;code&gt;wrangler deploy --temporary&lt;/code&gt; and no signup, live for 60 minutes and claimable afterward, and &lt;a href=&quot;https://vercel.com/changelog/vercel-functions-can-now-run-up-to-30-minutes?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Vercel Functions can now run up to 30 minutes&lt;/a&gt; for Pro and Enterprise teams on Node.js and Python, aimed at LLM reasoning, AI streaming, and document processing.&lt;/p&gt;
&lt;h3&gt;Tutorials&lt;/h3&gt;
&lt;ul&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/database/improve-query-performance-with-explain-plans-in-amazon-aurora-dsql?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Improve query performance with EXPLAIN plans in Amazon Aurora DSQL&lt;/a&gt; by Prema Iyer&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://ranthebuilder.cloud/blog/agentic-coding-hooks-deterministic-ai-guardrails?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Agentic Coding Hooks: Deterministic AI Guardrails&lt;/a&gt; by Ran Isenberg&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/machine-learning/shared-infrastructure-isolated-tenants-pool-model-multi-tenancy-with-amazon-bedrock-agentcore?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Shared infrastructure, isolated tenants: Pool model multi-tenancy with Amazon Bedrock AgentCore&lt;/a&gt; by Ashley Chen&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://dev.to/ntoledo319/migrating-aws-lambda-from-nodejs-20-to-22-every-breaking-change-8cp?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Migrating AWS Lambda from Node.js 20 to 22 — every breaking change&lt;/a&gt; by ntoledo319&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://letsdatascience.com/news/digitalocean-presents-hybrid-inference-pattern-for-ai-worklo-8a5fa047?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;DigitalOcean Presents Hybrid Inference Pattern for AI Workloads | Let&#39;s Data Science&lt;/a&gt; by&lt;/li&gt;
&lt;/ul&gt;
&lt;h3&gt;Reads&lt;/h3&gt;
&lt;p&gt;&lt;a href=&quot;https://dev.to/suletete/i-built-an-event-driven-order-system-with-both-ecs-and-lambda-heres-why-fcp?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;I built an event-driven order system with both ECS and Lambda. Here&#39;s why.&lt;/a&gt; by Suleiman Abdulkadir&lt;br /&gt;
Nice walkthrough of mixing ECS and Lambda instead of forcing everything into one compute model. The saga pattern with EventBridge is probably how I&#39;d build it too, though fifteen services for an order system is a lot of surface area to operate.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://devops.com/iac-isnt-dying-ai-makes-it-more-important?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;IaC Isn&#39;t Dying. AI Makes it More Important - DevOps.com&lt;/a&gt; by Jonah Kowall&lt;br /&gt;
The argument that IaC becomes your system of record for non-deterministic AI output is exactly right. If agents are generating infrastructure, you need something deterministic to reconcile against, and as of right now, that&#39;s still some form of IaC.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://anderson-mo-carvalho.medium.com/why-i-ripped-agentcore-and-strands-out-of-production-c3ed7551ec28?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Why I Ripped AgentCore and Strands Out of Production&lt;/a&gt; by Anderson Carvalho&lt;br /&gt;
This is the anti-framework story I keep seeing lately: the agent SDKs pile on abstraction you don&#39;t need until you suddenly do. Swapping back to Lambda-per-customer with direct Bedrock calls won&#39;t fit everyone, but matching complexity to the actual workload is the right instinct.&lt;/p&gt;
&lt;h3&gt;Podcasts, Videos, and more&lt;/h3&gt;
&lt;p&gt;&lt;a href=&quot;https://www.youtube.com/watch?v=PKG6dnt_VPA?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;What&#39;s new in Strands Agents | Serverless Office Hours&lt;/a&gt;&lt;br /&gt;
Julian Wood and the team run through the Strands updates, and Evals 1.0 is the one worth your time. Pre-production testing is the agent gap nobody has filled well yet.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://www.youtube.com/watch?v=Z5hdcD_Neco?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;The Great AI Reality Check Has Begun&lt;/a&gt;&lt;br /&gt;
The core point holds: generating code was never the hard part of software engineering, and 2025 drove that home. The &amp;quot;Doorman Fallacy&amp;quot; applies perfectly here, because the gap between code generation and shipping real systems gets really wide without the right people guiding it.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://www.youtube.com/watch?v=3ESclFr8m7I?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;MIT Just Revealed the AI Bubble&#39;s Fatal Flaw&lt;/a&gt;&lt;br /&gt;
The title oversells it, but the breakdown of who actually has the compute and data to compete is a useful gut check. Worth a watch if you want a clearer read on the economics underneath all the model announcements. AI will remain incredibly useful, but I don&#39;t think there&#39;s a moat that will sustain these valuations.&lt;/p&gt;
&lt;h3&gt;New from AWS&lt;/h3&gt;
&lt;ul&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/aws-transform-migrations-region-expansion?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS Transform for migrations now supports all AWS commercial regions as migration targets&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/aws-identity-center-separate-quotas?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS IAM Identity Center now supports separate quotas for AWS accounts and applications&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/amazon-msk-ai-agent-skills?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon MSK now offers AI Agent Skills to help developers operate MSK efficiently and accelerate migrations to MSK&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/amazon-cloudwatch-synthetics-multilocation?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon CloudWatch Synthetics now supports multilocation canaries&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/amazon-bedrock-agentcore-new-optimization-capabilities?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon Bedrock AgentCore introduces new optimization capabilities to continuously improve agents in production&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/amazon-mq-private-network-connectivity?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon MQ for RabbitMQ now supports private networking connectivity&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/amazon-sagemaker-ai-inference?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon SageMaker AI Announces New observability capability For Inference Endpoints&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/oracle-database-aws-autonomous-database-serverless?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Oracle Database@AWS now supports Oracle Autonomous AI Database Serverless&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/aws-glue-data-catalog?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS Glue Data Catalog now supports business context and semantic search (Preview)&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/safe-secrets-handling-in-agent-toolkit-for-aws?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS Secrets Manager introduces safe secrets handling in the Agent Toolkit for AWS&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/amazon-s3-annotations-business-context?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon S3 adds annotations to provide AI agents and analytics tools with context for data discovery&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/aws-security-agent-threat-modeling?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS Security Agent announces support for Threat Modeling&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/aws-kiro-power-claude-code?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS Security Agent adds Kiro Power, Claude Code, simulated validations and new integrations support&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/aws-transform-model-to-model-assessments?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS Transform now supports model-to-model migration assessment for generative AI workloads&lt;/a&gt;&lt;/li&gt;
&lt;/ul&gt;
&lt;h3&gt;Security&lt;/h3&gt;
&lt;p&gt;&lt;a href=&quot;https://gbhackers.com/hazybeacon-abuses-aws-lambda-function/amp?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;HazyBeacon Abuses AWS Lambda Function URLs for Stealthy Command-and-Control Operations&lt;/a&gt;&lt;br /&gt;
HazyBeacon uses stolen IAM credentials to stand up Lambda Function URLs as command-and-control channels that blend right into trusted AWS traffic. The takeaway is about identity governance and egress monitoring, since the technique leans on credential theft rather than any flaw in Lambda itself.&lt;/p&gt;
&lt;h3&gt;Events&lt;/h3&gt;
&lt;p&gt;&lt;strong&gt;June 29 - July 2, 2026 -&lt;/strong&gt; &lt;a href=&quot;https://ai.engineer/worldsfair?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23369&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AI Engineer World&#39;s Fair 2026: San Francisco&lt;/a&gt; 🗣️ (I&#39;ll be there!)&lt;/p&gt;
&lt;h3&gt;Final Thoughts 🤔&lt;/h3&gt;
&lt;p&gt;For as long as we&#39;ve been shipping to the cloud, there&#39;s been a wall between the code that does the work and the code that describes where it runs. You write a function, then you go write the CloudFormation, the Terraform, the CDK stack, or the SAM template that tells the cloud how to host it. Two artifacts, two mental models, kept in sync by hand and by hope. Infrastructure as Code was a real step forward because it made that second artifact deterministic and reviewable, and I&#39;m not here to argue against determinism. You want a system of record that says exactly what&#39;s running and why.&lt;/p&gt;
&lt;p&gt;What &lt;a href=&quot;https://docs.aws.amazon.com/blocks/latest/devguide/what-is-blocks.html&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS Blocks&lt;/a&gt; gets right is that the determinism doesn&#39;t have to live in a separate file. It can sit right next to the code that uses it. That&#39;s the same instinct behind the annotations Wing was doing, and the approaches others like Encore and Nitric have taken, where you declare what you need inside your application code and let the framework work out the provisioning. Watching AWS lean into that idea, smartly using TypeScript&#39;s integrated type safety, with a clean local-to-production story, is a good sign for where this is heading.&lt;/p&gt;
&lt;p&gt;Blocks doesn&#39;t go as far as I&#39;d like, and that gap is exactly the part I&#39;ve spent the last five-plus years on at Ampt. Mapping code to generated infrastructure is the easy half. The harder and more valuable half is making the infrastructure itself adaptable, a living thing that responds as the code and its usage patterns change, rather than a fixed target you have to keep reconciling and reshaping by hand. Blocks lines up closely with what we call &lt;a href=&quot;https://www.youtube.com/watch?v=-DZGnIhmOvE&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Productized Patterns&lt;/a&gt; at Ampt, and that adaptability is the direction I keep wanting more of.&lt;/p&gt;
&lt;p&gt;This matters more now than it did two years ago, because rapid code generation is the new normal. When code gets written this fast and a lot of it is throwaway, the old contract is backwards. Asking freshly generated code to also provision the infrastructure to run itself puts the burden in the wrong place. Flip it around: let the code be produced, and let the infrastructure figure out the best way to run it. That&#39;s a far better loop for testing and prototyping, and you can take the next step to harden it without locking yourself into something as rigid as a hand-maintained IaC stack.&lt;/p&gt;
&lt;p&gt;The conversation being back on the table, with AWS in it, is the real story here. Blocks is early and it&#39;s still in preview, but the principle underneath it is the one worth betting on: code you can throw away cheaply, and infrastructure that adapts to keep up. If the codegen tools are going to keep producing at this pace, that&#39;s the model that lets you move fast without leaving a pile of brittle stacks behind you.&lt;/p&gt;
&lt;p&gt;See you next week,&lt;br /&gt;
Jeremy&lt;/p&gt;
&lt;hr /&gt;
&lt;p&gt;I hope you enjoyed this newsletter. We&#39;re always looking for ideas and feedback to make it better and more inclusive, so please feel free to reach out to me via &lt;a href=&quot;https://bsky.app/profile/jeremydaly.com&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Bluesky&lt;/a&gt;, &lt;a href=&quot;http://www.linkedin.com/in/jeremydaly&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;LinkedIn&lt;/a&gt;, &lt;a href=&quot;http://twitter.com/jeremy_daly&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;X&lt;/a&gt;, or &lt;a href=&quot;mailto:contact@jeremydaly.com&quot;&gt;email&lt;/a&gt;.&lt;/p&gt;
</content>
  </entry>
  <entry>
    <title>Issue #368: Claude Fable 5 is currently unavailable 🚫</title>
    <link href="https://offbynone.io/issues/368/"/>
    <updated>2026-06-16T12:00:00Z</updated>
    <summary>In this issue, the US government orders Anthropic to pull Fable 5 and Mythos 5, AWS WAF starts charging AI bots for content, and Bedrock adds Grok, Gemma, and a pair of GPTs.</summary>
    <id>https://offbynone.io/issues/368/</id>
    <content type="html">&lt;h2&gt;Claude Fable 5 is currently unavailable 🚫&lt;/h2&gt;
&lt;p&gt;In our &lt;a href=&quot;https://offbynone.io/issues/367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;previous issue&lt;/a&gt;, Anthropic shipped two major models, DynamoDB got &amp;quot;extended&amp;quot; to run locally on Postgres, and Aurora DSQL added JSONB support. In this issue, the US government orders Anthropic to pull Fable 5 and Mythos 5, AWS WAF starts charging AI bots for content, and Bedrock adds Grok, Gemma, and a pair of GPTs. Plus, we&#39;ve got lots of great content from the cloud, serverless, and AI communities.&lt;/p&gt;
&lt;h3&gt;News &amp;amp; Announcements&lt;/h3&gt;
&lt;p&gt;The biggest story from last week is a takedown, not a product launch. &lt;a href=&quot;https://www.anthropic.com/news/fable-mythos-access?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Anthropic published a statement responding to a US government directive to suspend access to Fable 5 and Mythos 5&lt;/a&gt; over &amp;quot;national security concerns.&amp;quot; Anthropic walks through its defense-in-depth approach and argues the jailbreak vulnerabilities that triggered the order are about the same as the ones in models that are still happily serving traffic to North Korea (I may have made that last part up). Two weeks ago Fable 5 was the first generally available Mythos-class model on AWS. Now it&#39;s &lt;em&gt;gone&lt;/em&gt;. Whatever you think of the merits, watching a government switch off a frontier model overnight is a preview of a world many of us haven&#39;t planned for. The residency assumptions baked into your architecture diagrams may be softer than you think.&lt;/p&gt;
&lt;p&gt;Speaking of Bedrock, the model menu keeps growing. &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/grok-amazon-bedrock?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Grok 4.3 from xAI is now available on Amazon Bedrock&lt;/a&gt; with configurable reasoning effort levels, running on Mantle, the new inference engine AWS built for price performance. &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/gemma-4-amazon-bedrock?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Google DeepMind&#39;s Gemma 4 family landed too&lt;/a&gt;, three open-weight variants with reasoning, multimodal understanding across text, image, video, and audio, native function calling, 35+ languages, and 256K-token context windows (the &lt;a href=&quot;https://aws.amazon.com/blogs/machine-learning/introducing-gemma-4-models-on-amazon-bedrock?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS ML blog has the deeper writeup&lt;/a&gt; covering the bedrock-mantle endpoint and OpenAI-compatible APIs). And &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/openai-gpt-us-east-virginia-amazon?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;OpenAI&#39;s GPT-5.4 and GPT-5.5 are now in US East (N. Virginia)&lt;/a&gt;, both with 272K-token context and Responses API streaming, GPT-5.5 aimed at coding and research and GPT-5.4 at production reasoning. Three model families through one endpoint is great until the bill shows up, which is why the cost attribution work AWS has been doing will eventually pay off.&lt;/p&gt;
&lt;p&gt;The item I keep coming back to is the one that turns your bot traffic into a revenue line. &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/aws-waf-ai-traffic-monetization?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS WAF announced AI traffic monetization&lt;/a&gt; using the x402 protocol for machine-to-machine payments, letting publishers set differentiated pricing for AI bots and collect stablecoin payouts through Coinbase. The &lt;a href=&quot;https://aws.amazon.com/blogs/aws/aws-waf-adds-ai-traffic-monetization-capability-to-help-content-owners-charge-ai-bots-for-content-access?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS blog has the mechanics&lt;/a&gt;: WAF returns an HTTP 402 with a machine-readable JSON price manifest, works with CloudFront distributions, and settles through Coinbase&#39;s x402 Facilitator. It&#39;s not happening in a vacuum, either. &lt;a href=&quot;https://techcrunch.com/2026/05/28/visa-invests-in-replit-to-power-agentic-payments-for-developers?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Visa invested in Replit to power agentic payments for developers&lt;/a&gt;, including work on Visa&#39;s Trusted Agent Protocol, so the plumbing for agents that pay for things is getting built on multiple fronts.&lt;/p&gt;
&lt;p&gt;Agent platforms keep on maturing as well. &lt;a href=&quot;https://claude.com/blog/whats-new-in-claude-managed-agents?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Claude Managed Agents added scheduled deployments and environment vaults&lt;/a&gt;, with Rakuten and Notion already running recurring spreadsheet analysis and report generation, plus Browserbase and KERNEL integrations for browser work. &lt;a href=&quot;https://openai.com/index/openai-to-acquire-ona?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;OpenAI is acquiring Ona&lt;/a&gt; to give Codex persistent cloud execution, so agents can grind on a task for hours or days inside a customer-controlled environment. OpenAI also struck a deal to &lt;a href=&quot;https://openai.com/index/openai-on-oracle-cloud?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;let OCI customers reach its models and Codex through Oracle Universal Credits&lt;/a&gt;, wiring AI spend into existing enterprise purchasing. And &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/opensearch-agentic-observability-mcp-app?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon OpenSearch Service launched MCP Apps for agentic observability&lt;/a&gt;, letting agents dig into logs, traces, metrics, and alerts for root cause analysis from inside Claude Desktop or VS Code.&lt;/p&gt;
&lt;p&gt;On the data and ops side, &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/agentcore-memory-scmetadata?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon Bedrock AgentCore Memory now supports strictly consistent metadata for long-term memory&lt;/a&gt;, so you can attach values from your application that pass through without LLM inference. That gives you department-scoped retrieval, compliance boundaries, and multi-tenant memory where each tenant gets processed on its own, which is the kind of thing that sounds boring until you try to build memory for more than one customer. And &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/amazon-cross-account-metrics-centralization?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon CloudWatch added cross-account metrics centralization&lt;/a&gt; through AWS Organizations, replicating metrics from many accounts and regions into one destination account for unified monitoring and governance.&lt;/p&gt;
&lt;p&gt;A few more worth your attention. &lt;a href=&quot;https://aws.amazon.com/blogs/architecture/introducing-the-snowflake-and-aws-custom-lens-for-the-aws-well-architected-framework?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS and Snowflake released a joint Custom Lens for the Well-Architected Framework&lt;/a&gt;, folding both platforms&#39; best practices into one review across seven pillars, so you can stop juggling two separate sets of guidance. &lt;a href=&quot;https://aws.amazon.com/blogs/developer/aws-cli-v1-maintenance-mode-announcing-changes-to-dependency-updates?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS CLI v1 is entering maintenance mode in July 2026&lt;/a&gt;, with botocore and s3transfer vendored directly into the codebase, which means if you&#39;re running CLI v1 and boto3 side by side, they&#39;ll each carry their own copies from here on out. And &lt;a href=&quot;https://kiro.dev/blog/kiro-pro-max?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Kiro shipped a $100/mo Pro Max tier&lt;/a&gt; with more credits and access to all premium models. The jump from $40 to $200 was definitely a bit much for your average user, so dropping a tier right in the middle is a smart read of who actually churns.&lt;/p&gt;
&lt;p&gt;Finally, I shipped a new Prisma 7 adapter in the &lt;a href=&quot;https://www.jeremydaly.com/data-api-client-v2-4-prisma-support?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Data API Client v2.4&lt;/a&gt;, so now you can point Prisma, Knex, Drizzle, or Kysely at the RDS Data API for Provisioned or Serverless Aurora clusters without a connection pool or VPC.&lt;/p&gt;
&lt;h3&gt;Tutorials&lt;/h3&gt;
&lt;ul&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/machine-learning/ai-agent-failure-detection-and-root-cause-analysis-with-strands-evals?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AI Agent Failure Detection and Root Cause Analysis with Strands Evals&lt;/a&gt; by Po-Shin Chen&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/machine-learning/evaluate-ai-agents-systematically-with-agent-evalkit?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Evaluate AI agents systematically with Agent-EvalKit&lt;/a&gt; by Ishan Singh&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/architecture/how-samsung-achieved-real-time-pricing-with-aws-lambda-response-streaming?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;How Samsung achieved real-time pricing with AWS Lambda Response Streaming&lt;/a&gt; by Vijay Naik&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://dev.to/aws-heroes/serverless-applications-on-aws-with-lambda-using-java-25-api-gateway-and-aurora-dsql-lambda-4hbj?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Serverless applications on AWS with Lambda using Java 25, API Gateway and Aurora DSQL - Lambda performance optimization approaches&lt;/a&gt; by Vadym Kazulkin&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://dev.to/qasim157/run-your-email-agent-on-serverless-42d2?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Run Your Email Agent on Serverless&lt;/a&gt; by Qasim Muhammad&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://medium.com/@yogigupta79/the-death-of-tmp-s3-mounting-for-lambda-is-a-game-changer-f394d456be2c?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;The Death of /tmp: S3 Mounting for Lambda is a Game-Changer&lt;/a&gt; by Yogesh Gupta&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://medium.com/@chiragmehta900/cut-your-aws-fargate-bill-by-40-10-waste-patterns-i-fixed-in-production-d657d61469d2?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Cut Your AWS Fargate Bill by 40% — 10 Waste Patterns I Fixed in Production&lt;/a&gt; by Chirag Mehta&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://sodkiewiczm.medium.com/mcp-apps-because-your-users-deserve-more-than-a-wall-of-text-96e2fda9c9d9?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;MCP Apps: Because Your Users Deserve More Than a Wall of Text&lt;/a&gt; by Maciej Sodkiewicz&lt;/li&gt;
&lt;/ul&gt;
&lt;h3&gt;Reads&lt;/h3&gt;
&lt;p&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/machine-learning/how-frontier-teams-are-reinventing-ai-native-development?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;How frontier teams are reinventing AI-native development&lt;/a&gt;&lt;br /&gt;
Swami details three approaches AWS used to test AI-native workflows, including pathfinder initiatives and structured sprints, and lays out five practices for teams restructuring around autonomous agents. If you&#39;re still treating AI as a fancier autocomplete, this is a nudge to think bigger about how the work itself changes.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://dev.to/jverhoeks/the-review-bottleneck-rethinking-software-and-infrastructure-design-for-the-agent-era-752?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;The Review Bottleneck: Rethinking Software and Infrastructure Design for the Agent Era&lt;/a&gt;&lt;br /&gt;
A look at how coding agents moved the delivery bottleneck from writing code to reviewing and coordinating it. The proposed fixes, bounded contexts, contract-driven development, and pushing review upstream to intent instead of output, line up with what a lot of teams are feeling right now but haven&#39;t named yet.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://charity.wtf/2026/06/15/ai-demands-more-engineering-discipline-not-less-xpost?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AI demands more engineering discipline. Not less.&lt;/a&gt;&lt;br /&gt;
Charity Majors makes the case for what she calls Phoenix Architectures, where code becomes a materialized view you can regenerate once it goes stale. She draws the line from immutable infrastructure to treating AI-generated code as disposable, with validation moving to production. Classic Charity, and definitely worth your time.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://claude.com/blog/building-with-claude-managed-agents?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;The evolution of agentic surfaces: building with Claude Managed Agents&lt;/a&gt;&lt;br /&gt;
Anthropic introduces Claude Managed Agents as a set of composable APIs for production agents, handling orchestration, session management, credential isolation, and observability so teams can spend their time on context management instead of babysitting execution harnesses. Pairs well with the scheduling and vaults news above.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://dev.to/amitkayal/takeway-from-aws-generative-ai-lens-14dj?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Takeaways from AWS Generative AI Lens&lt;/a&gt;&lt;br /&gt;
Amit Kayal breaks down the AWS Generative AI Lens with a focus on controlled AI-assisted workflows versus fully autonomous agents, walking through when AI should classify, when it should recommend, and when it should actually execute. The data governance and multi-tenant sections are the parts I&#39;d read twice.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://serverlessdna.com/strands/lambda/vpc?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Lambda in a VPC Is Fine&lt;/a&gt;&lt;br /&gt;
Michael Walmsley walks through the evolution of Lambda VPC networking, from the painful 2016 days of on-demand ENI creation to today&#39;s Hyperplane implementation. If you&#39;re still repeating the old &amp;quot;never put Lambda in a VPC&amp;quot; advice, this explains why it stopped being true years ago.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://thenewstack.io/aws-opensearch-serverless-agentic-rebuild?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Why AWS scrapped OpenSearch&#39;s architecture to chase agent workloads&lt;/a&gt;&lt;br /&gt;
Frederic Lardinois of The New Stack covers AWS&#39;s near-complete rebuild of OpenSearch Serverless, with separated storage and compute that scales to zero when idle and auto-scales 20x faster than before. It&#39;s built for the burst-and-idle usage that agent workloads generate, with log analytics arriving in June and agent memory features in H2 2026.&lt;/p&gt;
&lt;h3&gt;New from AWS&lt;/h3&gt;
&lt;ul&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/amazon-route-53-resolver-dns?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon Route 53 Resolver DNS Firewall now supports Palo Alto Networks Advanced DNS Security (Preview)&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/amazon-cloudwatch-log-analytics?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon CloudWatch introduces Log Analytics for unified log analysis&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/aws-lambda-managed-instances-tag-propagation?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS Lambda Managed Instances now supports Tag Propagation for Managed Resources&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/aws-management-console-private?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS Management Console Private Access now works without internet connectivity&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/amazon-mwaa-serverless-eventbridge?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon MWAA Serverless now supports Amazon EventBridge notifications&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/amazon-sagemaker-ft-nemotron-3?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;SageMaker AI now supports serverless fine-tuning for NVIDIA Nemotron models&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/aws-cost-explorer?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS launches Cost Explorer historical data retention for accounts in billing groups&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/amazon-cloudwatch-query-studio-generally-available?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon CloudWatch Query Studio is now generally available&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/cloudwatch-application-signals-supports%20infrastructure-logs-traces-context-for-faster%20troubleshooting?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon CloudWatch Application Signals now supports infrastructure, logs, and traces context for faster troubleshooting&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/aws-cost-usage-report?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS Cost and Usage Report 2.0 now supports table configurations update&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/amazon-sagemaker-unified-studio-emr?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon SageMaker Unified Studio Notebooks now support EMR Serverless&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/aws-cli-agent-toolkit?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;The AWS Command Line Interface (CLI) now supports the Agent Toolkit for AWS&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/aws-workload-credentials-provider?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS announces AWS Workload Credentials Provider&lt;/a&gt;&lt;/li&gt;
&lt;/ul&gt;
&lt;h3&gt;Security&lt;/h3&gt;
&lt;p&gt;&lt;a href=&quot;https://securosis.com/blog/aws-destroyed-the-value-proposition-for-bedrock?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23368&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS Destroyed the Value Proposition for Bedrock&lt;/a&gt; by Chris Farris&lt;br /&gt;
Chris digs into the part of the Fable 5 and Mythos 5 launch nobody put in the headline: the only allowed retention mode for these models on Bedrock is &lt;code&gt;provider_data_share&lt;/code&gt;. Using them means your prompts and outputs leave the AWS boundary, land with Anthropic for 30 days, and become subject to human review. That breaks the neutral-broker guarantee that sent regulated and European shops to Bedrock in the first place. He walks through the compliance fallout and the SCP you should deploy today to deny anything other than &lt;code&gt;none&lt;/code&gt;. Read this before you point a workload at either model, assuming they get turned back on.&lt;/p&gt;
&lt;h3&gt;From Socials&lt;/h3&gt;
&lt;blockquote class=&quot;twitter-tweet&quot;&gt;&lt;p lang=&quot;en&quot; dir=&quot;ltr&quot;&gt;Just spent the last two weeks reworking my local Agent Hub system to use &lt;a href=&quot;https://x.com/opencode?ref_src=twsrc%5Etfw&quot;&gt;@opencode&lt;/a&gt; as the harness with qwen, gemma4, and mistral local models. Then I get this at 7:01pm. 😑 &lt;a href=&quot;https://t.co/T9as70aqdQ&quot;&gt;pic.twitter.com/T9as70aqdQ&lt;/a&gt;&lt;/p&gt;&amp;mdash; Jeremy Daly (@jeremy_daly) &lt;a href=&quot;https://x.com/jeremy_daly/status/2066673284888891725?ref_src=twsrc%5Etfw&quot;&gt;June 16, 2026&lt;/a&gt;&lt;/blockquote&gt; &lt;script async=&quot;&quot; src=&quot;https://platform.x.com/widgets.js&quot; charset=&quot;utf-8&quot;&gt;&lt;/script&gt;
I&#39;m not sure whether to be excited by this message, or if I should prepare for another rug pull. Either way, it forced me down an interesting multi-harness orchestration path.
&lt;h3&gt;Final Thoughts 🤔&lt;/h3&gt;
&lt;p&gt;HTTP 402 had been sitting in the spec since the early 90s with a note that said &amp;quot;reserved for future use.&amp;quot; For three decades it was the status code nobody got to use, a placeholder for a payment layer the web never seemed to materialize. Then about a year ago, Coinbase introduced &amp;quot;x402: An open standard for&lt;br /&gt;
internet-native payments.&amp;quot; Wait, did the crypto bros get it right? 😬 (fyi, I&#39;m still a hard no on that)&lt;/p&gt;
&lt;p&gt;AWS WAF now returns a 402 with a machine-readable price manifest when an AI bot asks for your content. The bot&#39;s agent reads the manifest, pays in stablecoin through Coinbase&#39;s x402 facilitator, and gets the content. No human in the loop and no checkout page. At the same time, Visa is putting money into Replit to build agentic payments and pushing its Trusted Agent Protocol, so the same machinery is getting assembled by the incumbents who actually move money for a living. When a 30-year-old dead status code and a Visa investment point in the same direction, that&#39;s usually a signal worth paying attention to.&lt;/p&gt;
&lt;p&gt;What&#39;s happening here is a shift in how we treat bots. For most of the web&#39;s history, automated traffic was something you blocked, rate-limited, or grudgingly tolerated. The robots.txt era assumed crawlers were either friendly enough to respect a text file or hostile enough to fight. Now there&#39;s a third option: charge them. If an agent wants your content badly enough to pay for it, you can let it, and you can put a number on exactly how much that access is worth.&lt;/p&gt;
&lt;p&gt;I&#39;m not sure this scales, and there are real reasons for skepticism. Stablecoin payouts assume a settlement story most finance teams haven&#39;t signed off on. Differentiated pricing for bots assumes agents will agree to pay instead of routing around you, and the whole thing has a chicken-and-egg problem where it only matters once enough agents speak the protocol and enough publishers demand payment. None of that is solved. But the direction is clear, and for the first time the economics of serving an AI bot aren&#39;t automatically negative.&lt;/p&gt;
&lt;p&gt;There&#39;s a question worth thinking about if you run content or an API. &amp;quot;Block all bots&amp;quot; is no longer the only defensive move available to you. The more interesting question is which agents you&#39;d actually want to charge, which ones you&#39;d serve for free because they send value back, and what your content is worth to a machine that has a budget and no patience for a paywall modal. That&#39;s a pricing exercise, not a security one, and most of us have never had to think about it. We probably should start.&lt;/p&gt;
&lt;p&gt;See you next week,&lt;br /&gt;
Jeremy&lt;/p&gt;
&lt;hr /&gt;
&lt;p&gt;I hope you enjoyed this newsletter. We&#39;re always looking for ideas and feedback to make it better and more inclusive, so please feel free to reach out to me via &lt;a href=&quot;https://bsky.app/profile/jeremydaly.com&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Bluesky&lt;/a&gt;, &lt;a href=&quot;http://www.linkedin.com/in/jeremydaly&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;LinkedIn&lt;/a&gt;, &lt;a href=&quot;http://twitter.com/jeremy_daly&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;X&lt;/a&gt;, or &lt;a href=&quot;mailto:contact@jeremydaly.com&quot;&gt;email&lt;/a&gt;.&lt;/p&gt;
</content>
  </entry>
  <entry>
    <title>Issue #367: What did I miss? 🎓</title>
    <link href="https://offbynone.io/issues/367/"/>
    <updated>2026-06-09T12:00:00Z</updated>
    <summary>In this issue, Anthropic ships two major models, DynamoDB gets &#39;extended&#39; to run locally on Postgres, and Aurora DSQL adds JSONB support.</summary>
    <id>https://offbynone.io/issues/367/</id>
    <content type="html">&lt;h2&gt;What did I miss? 🎓&lt;/h2&gt;
&lt;p&gt;I took a couple of weeks off, so we&#39;re playing catch-up. My youngest daughter graduated from high school last week, and between that, the after-prom party she threw at my house, and her graduation party (also at my house), there wasn&#39;t a lot of time left for keeping up with serverless, AI, and cloud. So this one covers about three weeks of news, and it&#39;s a long one. Apologies in advance.&lt;/p&gt;
&lt;p&gt;In this issue, Anthropic ships two major models, DynamoDB gets &amp;quot;extended&amp;quot; to run locally on Postgres, and Aurora DSQL adds JSONB support. Plus, we&#39;ve got plenty of awesome content from the cloud, serverless, and AI communities.&lt;/p&gt;
&lt;h3&gt;News &amp;amp; Announcements&lt;/h3&gt;
&lt;p&gt;Let&#39;s start with the money, because it&#39;s the reason for everything else. &lt;a href=&quot;https://www.anthropic.com/news/series-h?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Anthropic raised a boatload of money&lt;/a&gt;, $65 billion in a Series H at a $965 billion post-money valuation. That kind of capital buys a lot of compute, and the spending showed up almost immediately in the product line.&lt;/p&gt;
&lt;p&gt;First came &lt;a href=&quot;https://www.anthropic.com/news/claude-opus-4-8?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Claude Opus 4.8&lt;/a&gt;, which introduced dynamic workflows in Claude Code as a research preview, better coding and browser-automation numbers, and effort control settings, all at the same price as Opus 4.7. Then, before anyone had a chance to settle in, Anthropic announced &lt;a href=&quot;https://www.anthropic.com/news/claude-fable-5-mythos-5?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Claude Fable 5 and Claude Mythos 5&lt;/a&gt;, the first generation of Mythos-class models built for autonomous, professional work. Fable 5 is the one you can actually use, and Mythos 5 remains the locked-down sibling. If you want a second opinion before you commit, Claire Vo&#39;s &lt;a href=&quot;https://www.lennysnewsletter.com/p/claude-fable-5-review-what-the-new?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;review of Fable 5&lt;/a&gt; puts it through three real-world scenarios and is honest about where it falls down.&lt;/p&gt;
&lt;p&gt;AWS, predictably, did not want to be left out. &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/claude-opus-4.8-aws?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Claude Opus 4.8 landed on AWS&lt;/a&gt; through Bedrock and Claude Platform, and then &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/claude-fable-5-aws?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Fable 5 showed up as the first generally available Mythos-class model on AWS&lt;/a&gt; too, with a &lt;a href=&quot;https://aws.amazon.com/blogs/aws/anthropic-claude-fable-5-on-aws-mythos-class-capabilities-with-built-in-safeguards-now-available?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;longer writeup on the AWS blog&lt;/a&gt; covering the built-in safeguards for autonomous operation. Anthropic wasn&#39;t the only model vendor getting the Bedrock treatment, either. &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/amazon-bedrock-openai-models-codex-generally-available?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;OpenAI&#39;s GPT-5.5, GPT-5.4, and Codex are now generally available on Bedrock&lt;/a&gt; with &lt;a href=&quot;https://aws.amazon.com/blogs/machine-learning/openai-models-and-codex-on-amazon-bedrock-are-now-generally-available?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;pay-per-token pricing matching OpenAI&#39;s direct rates&lt;/a&gt;, inference staying inside your chosen region, and the usual KMS, VPC, and CloudTrail story for compliance.&lt;/p&gt;
&lt;p&gt;To make all of this easier to work with, Bedrock also shipped a &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/amazon-bedrock-redesigned-console-optimized-openai-anthropic-compatible-apis?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;redesigned console optimized for the OpenAI- and Anthropic-compatible APIs&lt;/a&gt; (there&#39;s a &lt;a href=&quot;https://aws.amazon.com/blogs/aws/try-the-new-console-experience-in-amazon-bedrock-optimized-for-anthropic-and-openai-compatible-apis?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;hands-on writeup on the AWS blog&lt;/a&gt;) built around the &lt;code&gt;bedrock-mantle&lt;/code&gt; endpoint, with project-based organization, side-by-side comparisons, and prefilled code snippets. They rounded it out with &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/amazon-bedrock-request-level-usage-attribution?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;request-level usage attribution&lt;/a&gt; so you can tag individual inference calls by team or environment, &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/amazon-bedrock-supports-cloudwatch-metrics-bedrock-mantle-endpoint?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;CloudWatch metrics for the mantle endpoint&lt;/a&gt;, and &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/5/amazon-bedrock-service-quotas?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;expanded Service Quotas support&lt;/a&gt;. The cost attribution piece is the one I&#39;d pay attention to. Once you&#39;ve got three model families running through one endpoint, knowing which team is spending what stops being optional.&lt;/p&gt;
&lt;p&gt;The agent side of Bedrock kept pace. &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/amazon-bedrock-agentcore-runtime?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AgentCore Runtime added interactive shells&lt;/a&gt; via a new &lt;code&gt;InvokeAgentRuntimeCommandShell&lt;/code&gt; API, giving you WebSocket terminal access into a running agent&#39;s microVM to inspect files, run commands, or debug state without losing session context. &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/agentcore-identity-secrets-manager?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AgentCore Identity now lets you bring your own secrets through AWS Secrets Manager&lt;/a&gt;, and &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/aws-step-functions-agentcore?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Step Functions added an AgentCore-powered agentic reasoning step&lt;/a&gt; so you can drop a reasoning task into a state machine without bolting on extra infrastructure. The &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/aws-mcp-server?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS MCP Server picked up cross-account and cross-role access&lt;/a&gt; too, so a coding agent can finally hop between accounts and roles in a single session instead of stopping, swapping credentials, and starting over. Anyone who&#39;s managed agents across more than one account knows exactly how annoying that loop was.&lt;/p&gt;
&lt;p&gt;The most interesting database news of the bunch didn&#39;t get a flashy launch event. &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/aws-extenddb-dynamodb?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS released ExtendDB 0.1&lt;/a&gt;, an open source adapter that implements the DynamoDB API on top of pluggable storage backends, with PostgreSQL as the first reference implementation. That means you can write code against DynamoDB programming patterns and run it locally, in CI, or on-prem against Postgres. I&#39;ve been wanting something like this for years. DynamoDB Local has always been a reasonable stand-in, but a pluggable adapter that lets you point real DynamoDB access patterns at a Postgres backend opens up a lot of testing and migration scenarios that used to be a pain. It&#39;s 0.1, so temper your expectations, but the direction is genuinely useful.&lt;/p&gt;
&lt;p&gt;Aurora DSQL stayed busy, picking up &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/amazon-aurora-dsql-supports-jsonb?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;JSONB support with compression on by default&lt;/a&gt;, so you can store semi-structured config and API parameters next to your relational data and let DSQL compress the larger payloads for you. Over in search, the &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/amazon-opensearch-serverless-next-generation-generally-available?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;next generation of Amazon OpenSearch Serverless went GA&lt;/a&gt;, and the headline feature is scale-to-zero. There&#39;s a &lt;a href=&quot;https://aws.amazon.com/blogs/aws/introducing-the-next-generation-of-amazon-opensearch-serverless-for-building-your-agentic-ai-applications?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;proper deep-dive on the AWS blog&lt;/a&gt; that leans into the agentic AI angle with instant resource creation and Vercel and Kiro integrations, and OpenSearch Serverless also &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/opensearch-agentic-search?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;added Agentic Search&lt;/a&gt; on top. Scale-to-zero is the big one for me. Vector and search backends that scale to zero change the math on a whole category of side projects and low-traffic workloads that previously couldn&#39;t justify the always-on cost.&lt;/p&gt;
&lt;p&gt;A small but welcome bit of housekeeping: AWS is &lt;a href=&quot;https://aws.amazon.com/blogs/developer/announcing-updated-retry-behavior-for-aws-sdks-and-tools?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;standardizing retry behavior across all SDKs and tools&lt;/a&gt;. The change splits backoff into two strategies, a fast 50ms for transient errors and a slower 1000ms for throttling, which is a more sensible default than treating every failure the same way. It becomes the default in November 2026, but you can opt in today with &lt;code&gt;AWS_NEW_RETRIES_2026=true&lt;/code&gt;. If you&#39;ve ever hand-tuned retry configs to stop hammering a throttled service, this is the kind of quiet fix that saves you from rediscovering the same lesson on the next project.&lt;/p&gt;
&lt;p&gt;There was plenty more from AWS over the past few weeks. &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/aws-finops-agent-preview?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;FinOps Agent went into preview&lt;/a&gt;, answering cost questions and surfacing optimization opportunities out of Cost Optimization Hub and Compute Optimizer. &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/amazon-cognito-multi-region?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Cognito added multi-Region replication&lt;/a&gt; as an add-on for Essentials and Plus tier user pools, syncing identities to a standby Region so you can redirect traffic during a regional disruption. And AWS &lt;a href=&quot;https://aws.amazon.com/blogs/aws/meet-our-newest-aws-heroes-may-2026?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;named four new Heroes for May 2026&lt;/a&gt;, with serverless and AI/ML leaders from Italy, Canada, and Argentina. Congratulations to all of them. The community is better for the work you do.&lt;/p&gt;
&lt;p&gt;One last thing from me. I pushed an update to &lt;a href=&quot;https://github.com/jeremydaly/data-api-client?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;data-api-client&lt;/a&gt;, my &lt;code&gt;DocumentClient&lt;/code&gt;-style wrapper for the Amazon Aurora Serverless Data API. If you&#39;re working with the Data API and want the familiar parameter-mapping ergonomics instead of the raw request format, give it a look.&lt;/p&gt;
&lt;h3&gt;Tutorials&lt;/h3&gt;
&lt;ul&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/database/building-type-safe-applications-with-drizzle-orm-in-aurora-dsql?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Building type-safe applications with Drizzle ORM in Aurora DSQL&lt;/a&gt; by Dipen Patel&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/database/pagination-patterns-in-amazon-aurora-dsql?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Pagination patterns in Amazon Aurora DSQL&lt;/a&gt; by Sandhya Khanderia&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/machine-learning/its-safe-to-close-your-laptop-now-hosting-coding-agents-on-amazon-bedrock-agentcore?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;It&#39;s safe to close your laptop now: Hosting coding agents on Amazon Bedrock AgentCore&lt;/a&gt; by Evandro Franco&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/machine-learning/break-the-context-window-barrier-with-amazon-bedrock-agentcore?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Break the context window barrier with Amazon Bedrock AgentCore&lt;/a&gt; by Yuan Tian&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/machine-learning/building-multi-tenant-agents-with-amazon-bedrock-agentcore?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Building multi-tenant agents with Amazon Bedrock AgentCore&lt;/a&gt; by Dhawalkumar Patel&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/database/best-practices-for-amazon-dynamodb-global-tables-part-1-operational-readiness?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Best practices for Amazon DynamoDB Global Tables – Part 1: Operational readiness&lt;/a&gt; by Lee Hannigan&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/database/best-practices-for-amazon-dynamodb-global-tables-part-2-failover-strategies?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Best practices for Amazon DynamoDB Global Tables – Part 2: Failover strategies&lt;/a&gt; by Lee Hannigan&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/database/best-practices-for-amazon-dynamodb-global-tables-part-3-validating-regional-resilience-with-aws-fault-injection-service?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Best practices for Amazon DynamoDB Global Tables – Part 3: Validating regional resilience with AWS Fault Injection Service&lt;/a&gt; by Lee Hannigan&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://gunnargrosch.com/posts/sms-delivery-receipts-on-aws-lambda?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;SMS Delivery Receipts on AWS Lambda&lt;/a&gt; by Gunnar Grosch&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://www.readysetcloud.io/blog/allen.helton/your-agent-is-repeating-itself?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Your agent is repeating itself&lt;/a&gt; by Allen Helton&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://rustysl.com/en/blog/s3-on-demand-archive?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;On-Demand Archives on S3&lt;/a&gt; by Jérémie Rodon&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://darryl-ruggles.cloud/live-canary-deployments-with-aws-sam-the-new-websocket-api-resource-and-lambda-durable-functions?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS SAM WebSocket &amp;amp; Lambda Durable Functions: Canary Deploy&lt;/a&gt; by Darryl Ruggles&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://dev.to/aws-heroes/serverless-applications-on-aws-with-lambda-using-java-25-api-gateway-and-dynamodb-part-7-lambda-4po1?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Serverless applications on AWS with Lambda using Java 25, API Gateway and DynamoDB - Part 7 Lambda performance optimization approaches&lt;/a&gt; by Vadym Kazulkin&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://dev.to/aws-heroes/aws-lambda-managed-instances-with-java-25-and-aws-sam-part-7-implement-scheduled-scaling-4df9?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS Lambda Managed Instances with Java 25 and AWS SAM – Part 7 Implement scheduled scaling&lt;/a&gt; by Vadym Kazulkin&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://dev.to/aws-builders/s3-files-killed-my-least-favorite-lambda-pattern-25f9?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;S3 Files Killed My Least Favorite Lambda Pattern&lt;/a&gt; by Mwanza Simi&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://heeki.medium.com/closing-the-loop-from-code-generation-to-sandboxed-code-execution-b90cff6adbf7?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Closing the loop from code generation to sandboxed code execution&lt;/a&gt; by Heeki Park&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://pubudu.dev/posts/lambda-durable-functions-triggered-by-sqs?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Triggering Lambda Durable Functions from SQS&lt;/a&gt; by Pubudu Jayawardana&lt;/li&gt;
&lt;/ul&gt;
&lt;h3&gt;Reads&lt;/h3&gt;
&lt;p&gt;&lt;a href=&quot;https://darryl-ruggles.cloud/the-real-cost-of-vector-storage-s3-vectors-vs-opensearch-vs-pgvector-vs-pinecone?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Vector Storage Costs: S3, OpenSearch, pgvector, Pinecone&lt;/a&gt; by Darryl Ruggles&lt;br /&gt;
Darryl built a full cost model and benchmark harness comparing S3 Vectors, OpenSearch Serverless NextGen, Aurora pgvector, and Pinecone, including how the May 2026 scale-to-zero launch shifts the comparison. There&#39;s a calculator to find the crossover point for your own workload shape, which is exactly the kind of thing you want before you pick a vector store and regret it later.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://ranthebuilder.cloud/blog/ai-changed-how-we-build-our-tools-didn-t?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AI Changed How We Build. Our Tools Didn&#39;t.&lt;/a&gt; by Ran Isenberg&lt;br /&gt;
Ran walks through the gap between AI-driven development and the tooling we still use to manage it. IDEs, GitHub, Jira, and sprint planning were all built for a world where humans wrote the code, and they haven&#39;t caught up to one where agents write and engineers mostly review. He&#39;s got a &lt;a href=&quot;https://ranthebuilder.cloud/blog/ai-changed-the-engineer-s-job-here-s-how-to-adapt?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;companion piece on adapting the engineer&#39;s job&lt;/a&gt; that gets into burnout risk and rising token costs too.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://charity.wtf/2026/06/02/ai-enthusiasts-are-in-a-race-against-time-ai-skeptics-are-in-a-race-against-entropy-xpost?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AI enthusiasts are in a race against time, AI skeptics are in a race against entropy&lt;/a&gt; by Charity Majors&lt;br /&gt;
Charity uses Fin&#39;s productivity gains as a case study and lands on a point that&#39;s easy to lose in the hype: the wins came from engineering discipline and fast feedback loops, not from AI being magic. If you&#39;re trying to bridge the gap between the true believers and the people rolling their eyes on your team, this is a good framework.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://claude.com/blog/running-an-ai-native-engineering-org?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Running an AI-native engineering org&lt;/a&gt;&lt;br /&gt;
Anthropic&#39;s engineering team shares how their process changed once every commit became Claude-assisted, including the move from six-month roadmaps to far more fluid planning. The bit about going past &amp;quot;who changed this&amp;quot; to &amp;quot;what information do I actually need&amp;quot; is the part worth sitting with.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://claude.com/blog/lessons-from-building-claude-code-how-we-use-skills?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Lessons from building Claude Code: How we use skills&lt;/a&gt;&lt;br /&gt;
The Claude Code team breaks down nine skill types they use internally, from library reference to verification to scaffolding, plus the practices that make a skill actually work. If you&#39;re building anything with skills, this is required reading.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://claude.com/blog/using-claude-code-the-unreasonable-effectiveness-of-html?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Using Claude Code: The unreasonable effectiveness of HTML&lt;/a&gt; by Thariq Shihipar&lt;br /&gt;
Anthropic makes the case that HTML beats Markdown for AI output because of its density and interactivity, with examples spanning richer docs, code reviews, and throwaway custom editors. I said it last issue and I&#39;ll say it again: I&#39;m sold on the HTML move.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://serverlessdna.com/strands/ai-agents/agent-loops-are-hungrier-than-you-think?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Your Agent Loops are Hungrier Than You Think&lt;/a&gt; by Michael Walmsley&lt;br /&gt;
Michael lays out why agentic loops burn tokens quadratically: every turn replays the full conversation history, so turn 20 is paying for turns 1 through 19. He backs it with real token counts from actual scenarios, and if you&#39;ve been surprised by an agent bill, this explains where it went.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://claude.com/blog/the-claude-cowork-product-guide?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;The Claude Cowork product guide&lt;/a&gt;&lt;br /&gt;
Anthropic&#39;s guide to Claude Cowork, their desktop knowledge-work agent, covers local file access, Slack and Google Drive integration, when to reach for it over other Claude tools, and seven worked examples. A useful orientation if you&#39;re trying to figure out where Cowork fits.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://openai.com/index/codex-for-knowledge-work?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Codex is becoming a productivity tool for everyone&lt;/a&gt;&lt;br /&gt;
OpenAI shared usage data putting Codex at 5 million weekly active users, with knowledge workers growing three times faster than developers. The use cases have spread well past code into reports, spreadsheets, presentations, and analysis. The line between &amp;quot;coding tool&amp;quot; and &amp;quot;work tool&amp;quot; keeps getting blurrier.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://blogs.oracle.com/developers/from-rag-to-memory-systems-building-stateful-ai-architecture?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AI Memory Systems Explained: From Retrieval to Durable, Context-Aware Agents&lt;/a&gt; by Jeremy Daly&lt;br /&gt;
This is mine. It&#39;s a deep architectural walkthrough of how to move from basic RAG to a production-grade memory system, covering five memory types (policy, preference, fact, episodic, trace), how their storage patterns differ, hybrid retrieval, and why you need a memory manager controlling what gets stored and retrieved while keeping governance and privacy intact. If memory has been the fuzzy part of your agent design, this should sharpen it up.&lt;/p&gt;
&lt;h3&gt;Podcasts, Videos, and more&lt;/h3&gt;
&lt;p&gt;&lt;a href=&quot;https://www.youtube.com/watch?v=5KWIf0mFzy8?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Building TypeScript agents with Strands | Serverless Office Hours&lt;/a&gt;&lt;br /&gt;
Erik walks through the Strands Agents TypeScript SDK for building agents on AWS, including agents that run in Node.js and the browser, connecting multiple model providers, and orchestrating multi-agent workflows.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://www.youtube.com/watch?v=RzwYFL6wIOU?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Building with Claude: Lessons from real projects | Serverless Office Hours&lt;/a&gt;&lt;br /&gt;
Ran Isenberg joins Julian Wood to talk through practical Claude Code workflows in serverless development: custom skills, configuration strategies, and context management. Worth watching if you&#39;re still figuring out how these tools fit your process.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://www.youtube.com/watch?v=R5cA74Qv4hs?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AI-assisted development in practice | Serverless Office Hours&lt;/a&gt;&lt;br /&gt;
Darryl Ruggles builds a full serverless blogging platform with AI coding tools and is honest about what works (MCP servers for Terraform and AWS docs), what breaks, and how to keep security and best practices intact when you let AI write your infrastructure.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://www.youtube.com/watch?v=dWIng6-7wQw?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Serverless Craic Ep86 AI and Software Development - the Real Problem&lt;/a&gt;&lt;br /&gt;
The Serverless Edge crew makes the case that AI amplifies both good and bad engineering practices, with a discussion that wanders through platform engineering, cognitive load, and socio-technical systems.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://www.youtube.com/watch?v=kquxaSSj3BE?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Serverless CrAIc Ep85 Why Team Topologies Matters More Than Ever in the AI Era&lt;/a&gt;&lt;br /&gt;
The crew asks whether AI agents count as team members and what that does to cognitive load, working through how organizational frameworks bend when code generation speeds up but human collaboration stays the bottleneck.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://www.youtube.com/watch?v=lqw92QYqM88?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS Bites #154: S3 Files&lt;/a&gt;&lt;br /&gt;
Eoin and Luciano dig into S3 Files, explaining why S3 was never really a file system (no atomic renames, expensive listings, immutable objects) and how this service bridges the gap, with benchmark data and a frank look at the 60-second write-back delay and eventual consistency.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://www.lennysnewsletter.com/p/claude-fable-5-review-what-the-new?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Claude Fable 5 review: what the new Mythos model gets right (and very wrong)&lt;/a&gt;&lt;br /&gt;
Claire Vo reviews Anthropic&#39;s first generally available Mythos-class model and the launches around it, including Managed Agents and safety classifiers, testing it on product specs and multi-agent orchestration. A grounded look at a model that&#39;s getting a lot of breathless coverage everywhere else.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://www.lennysnewsletter.com/p/a-rational-conversation-on-where?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;A rational conversation on where AI is actually going | Benedict Evans&lt;/a&gt;&lt;br /&gt;
Benedict Evans argues foundation models won&#39;t hold lasting pricing power and that value moves up the stack, with distribution becoming the real moat now that software is cheap to build. A nice counterweight to the model-vendor news in this issue.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://www.lennysnewsletter.com/p/the-ai-paradox-dan-shipper?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;The AI paradox: More automation, more humans, more work | Dan Shipper&lt;/a&gt;&lt;br /&gt;
Dan Shipper draws on running Every to argue that work is moving inside AI agents, that SaaS is thriving rather than dying because agents drive more usage, and that roles like PM are getting more leverage from AI tooling.&lt;/p&gt;
&lt;h3&gt;New from AWS&lt;/h3&gt;
&lt;ul&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/secrets-manager-managed-external-secrets-datadog-snowflake?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS Secrets Manager adds managed external secrets support for Datadog vended keys and Snowflake Programmatic Access Tokens&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/aws-security-agent?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS Security Agent adds verification scripts for pentest findings&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/amazon-ecs-pause-continue-deployments?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon ECS introduces pause and continue controls for service deployments&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/5/docdb8-serverless?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon DocumentDB (with MongoDB compatibility) Serverless is now available on DocumentDB 8.0&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/amazon-cloudwatch-logs-insights-new?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon CloudWatch Logs Insights adds 23 new query commands and functions&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/aws-lambda-managed-instances-region-expansion?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS Lambda Managed Instances expands to additional AWS Regions&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/amazon-emr-serverless-spark-connect?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Run Interactive Workloads on Amazon EMR Serverless with Spark Connect&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/aws-msk-express-topic-support-kstreams?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon MSK Express Brokers now support automatic topic creation with Kafka Streams&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/aws-cost-explorer-intelligent-cost-explanations?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS Cost Explorer launches intelligent cost explanations powered by Amazon Q&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/aws-ai-powered-cost-investigations?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS now provides AI-powered cost investigations for cost anomalies&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/amazon-redshift-incremental-manual-snapshots?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon Redshift reduces manual snapshot cost for Serverless and RG instances&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/aws-savings-plans-coverage?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS Savings Plans Purchase Analyzer now supports target coverage analysis&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/amazon-cloudwatch-mi-extended-retention-region-expansion?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon CloudWatch now supports querying metrics data up to two weeks old&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/durability-amazon-elasticache?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon ElastiCache for Valkey now supports durability&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/monitor-aws-budgets-using-dashboards?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Monitor AWS Budgets directly in Billing and Cost Management Dashboards with new Budgets widget&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/oracle-database-aws-available-twenty-regions?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Oracle Database@AWS is now available in twenty AWS Regions&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/amazon-ses-global-deliverability?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon SES now offers inbox placement metrics and blocklist monitoring&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/amazon-ses-tenant-level-suppression-lists?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon SES now supports tenant-level suppression lists&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/aws-shield-ddos?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS Shield Advanced introduces DDoS attack flow logs&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/06/keyspaces-cdc-iterator-position?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon Keyspaces (for Apache Cassandra) now provides CDC iterator position&lt;/a&gt;&lt;/li&gt;
&lt;/ul&gt;
&lt;h3&gt;Developer Tools&lt;/h3&gt;
&lt;p&gt;&lt;a href=&quot;https://dynamosql.com/?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;DynamoSQL™ — ANSI SQL for Amazon DynamoDB&lt;/a&gt;&lt;br /&gt;
DynamoSQL is a SQL query engine for DynamoDB with JOINs, CTEs, aggregations, and subqueries, no pipelines or ETL required. It&#39;s in beta with early access through AWS Marketplace and offers MCP integration for AI applications.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://dev.to/himaan4149/i-built-pretext-pdf-serverless-pdfs-without-chromium-1pg5?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;I Built pretext-pdf: Serverless PDFs Without Chromium&lt;/a&gt; by Himanshu Jain&lt;br /&gt;
Himanshu built pretext-pdf, a Node.js library that generates PDFs from JSON without Chromium, aimed at structured documents like invoices and reports with 40-100ms generation times. If you&#39;ve ever wrestled a headless Chromium into a Lambda just to make a PDF, this is a lighter path.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/developer/introducing-open-source-skills-for-aws-sdk-best-practices?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23367&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Introducing Open-Source Skills for AWS SDK Best Practices&lt;/a&gt; by David Yaffe&lt;br /&gt;
AWS released open-source skills for their Agent Toolkit to improve how AI coding agents generate SDK code, currently for Swift, JavaScript v3, and Python (Boto3), targeting the common mistakes like wrong API names, bad parameter types, and missed paginators.&lt;/p&gt;
&lt;h3&gt;Final Thoughts 🤔&lt;/h3&gt;
&lt;p&gt;That $65 billion raise is the smoke rising from the Anthropic and OpenAI IPO talk, with their valuations looking shakier than the headlines suggest once you do the math on token economics. Burning compute to win benchmarks is one thing. Making the unit economics work when customers actually use the product is another, and that&#39;s where the recent billing changes come in. Anthropic pulling &lt;code&gt;claude -p&lt;/code&gt; out of what your Max subscription covers, plus the GitHub Copilot billing changes are already having a real effect on how people use these tools. The tokenmaxing that let everyone ship slop faster is getting expensive, and maybe that&#39;s (kind of) a good thing.&lt;/p&gt;
&lt;p&gt;It forces discipline, which is the thread running through several pieces in this issue. Charity Majors makes the case that the AI productivity wins came from engineering discipline and tight feedback loops, not magic. Ran Isenberg points out that our tools were built for humans writing code and are straining under agents doing it. Both are circling the same idea: the teams that come out ahead won&#39;t be the ones with the most tokens, they&#39;ll be the ones with the most discipline. If that discipline doesn&#39;t show up, we&#39;re all in trouble.&lt;/p&gt;
&lt;p&gt;The model you use this year will be obsolete by next. The patterns you build around storage, cost, and testing will outlast all of them.&lt;/p&gt;
&lt;p&gt;See you next week,&lt;br /&gt;
Jeremy&lt;/p&gt;
&lt;hr /&gt;
&lt;p&gt;I hope you enjoyed this newsletter. We&#39;re always looking for ideas and feedback to make it better and more inclusive, so please feel free to reach out to me via &lt;a href=&quot;https://bsky.app/profile/jeremydaly.com&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Bluesky&lt;/a&gt;, &lt;a href=&quot;http://www.linkedin.com/in/jeremydaly&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;LinkedIn&lt;/a&gt;, &lt;a href=&quot;http://twitter.com/jeremy_daly&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;X&lt;/a&gt;, or &lt;a href=&quot;mailto:contact@jeremydaly.com&quot;&gt;email&lt;/a&gt;.&lt;/p&gt;
</content>
  </entry>
  <entry>
    <title>Issue #366: The Flat-Rate Honeymoon is Over 📈</title>
    <link href="https://offbynone.io/issues/366/"/>
    <updated>2026-05-19T12:00:00Z</updated>
    <summary>In this issue, Anthropic brings subscription clarity to Claude, Codex goes mobile, and Amazon DSQL gets CDC on DPUs.</summary>
    <id>https://offbynone.io/issues/366/</id>
    <content type="html">&lt;h2&gt;The Flat-Rate Honeymoon is Over 📈&lt;/h2&gt;
&lt;p&gt;In our &lt;a href=&quot;https://offbynone.io/issues/365&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;previous issue&lt;/a&gt;, Claude Platform set up shop on AWS, ElastiCache learned to do full-text and hybrid search, and Ampt rolled out Node.js 24 as the default runtime. This week, Anthropic brings subscription clarity to Claude, Codex goes mobile, and Amazon DSQL gets CDC on DPUs. Plus, we&#39;ve got plenty of awesome content from cloud, serverless, and AI communities.&lt;/p&gt;
&lt;h3&gt;News &amp;amp; Announcements&lt;/h3&gt;
&lt;p&gt;Anthropic announced this past week that &lt;a href=&quot;https://x.com/ClaudeDevs/status/2054610152817619388?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;paid Claude plans will get a dedicated monthly credit for programmatic usage starting June 15&lt;/a&gt;. Pro gets $20 in monthly credits, Max 20x gets $200, and anything past the allocation rolls onto API rates. If you&#39;ve been running SDK loops, &lt;code&gt;claude -p&lt;/code&gt; jobs, or GitHub Actions agents on your subscription, the honeymoon is over. Theo Browne took it about as well as you&#39;d expect. His &lt;a href=&quot;https://www.youtube.com/watch?v=131yAOjxHHQ?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;reaction video&lt;/a&gt; is appropriately titled &amp;quot;I&#39;m done,&amp;quot; and his X post promised to &lt;a href=&quot;https://x.com/theo/status/2054734057368621176?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;donate $10 to open source for every screenshot of a cancelled Claude Code plan&lt;/a&gt; that was shared. He&#39;s not wrong to be frustrated. The &lt;a href=&quot;https://x.com/mattpocockuk/status/2040536403289764275&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;rules have been ambiguous for quite some time&lt;/a&gt;, and this announcement does provide some clarity, just not what most of us were hoping for. The bigger lesson here is about AI platform risk. Cancelling Claude Code doesn&#39;t fix the real problem, and if your business depends on one vendor&#39;s pricing staying frozen forever, your subscription isn&#39;t the thing that needs changing.&lt;/p&gt;
&lt;p&gt;Even with all the developer backlash, Anthropic seems undeterred. They continue to ship more and more enterprise plumbing. &lt;a href=&quot;https://claude.com/blog/claude-managed-agents-updates?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Claude Managed Agents now support self-hosted sandboxes and MCP tunnels&lt;/a&gt;, so your agents&#39; tools can run inside your own infrastructure while orchestration stays on Anthropic&#39;s platform. Cloudflare, Daytona, Modal, and Vercel are all in the launch lineup, with &lt;a href=&quot;https://blog.cloudflare.com/claude-managed-agents?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Cloudflare getting its own first-class slot&lt;/a&gt; including Browser Run for automation and quick-start templates. Anthropic also pushed two vertical packages: &lt;a href=&quot;https://claude.com/blog/claude-for-the-legal-industry?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Claude for the legal industry&lt;/a&gt; and &lt;a href=&quot;https://www.anthropic.com/news/claude-for-small-business?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Claude for Small Business&lt;/a&gt;. Legal feels like the obvious play. High-value industries with messy document workflows are where managed agents make a lot of sense, so long as it doesn&#39;t keep hallucinating case law.&lt;/p&gt;
&lt;p&gt;Elsewhere in the agent space, OpenAI brought &lt;a href=&quot;https://openai.com/index/work-with-codex-from-anywhere?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Codex to the ChatGPT mobile app&lt;/a&gt; with Remote SSH now GA, programmatic access tokens, and HIPAA compliance for healthcare. Coding from your phone still sounds like a stretch, but the strategy is right: meet developers wherever they happen to be. Also, &lt;a href=&quot;https://thenewstack.io/temporal-replay-2026-news?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Temporal added Workflow Streams and Standalone Activities to its durable execution platform&lt;/a&gt;, both aimed squarely at the AI-in-production folks. Durability and debuggability are the two things most agentic systems are starving for, so the direction makes sense.&lt;/p&gt;
&lt;p&gt;AWS had a busy week too. &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/amazon-aurora-dsql-change-data-capture-preview?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon Aurora DSQL now supports change data capture in preview&lt;/a&gt;, streaming database changes to Kinesis Data Streams for event-driven apps and real-time analytics. I think the Distributed Processing Units plus Kinesis pricing is going to trip a few people up, but I&#39;ve been wrong before. &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/amazon-bedrock-advanced-prompt-optimization-migration-tool?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon Bedrock launched Advanced Prompt Optimization&lt;/a&gt; (with a &lt;a href=&quot;https://aws.amazon.com/blogs/aws/amazon-bedrock-introduces-new-advanced-prompt-optimization-and-migration-tool?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;deeper writeup on the AWS blog&lt;/a&gt;), automating prompt comparison across up to 5 models with custom evaluation metrics, Lambda-based scoring, LLM-as-a-Judge rubrics, and multimodal inputs including images and PDFs.&lt;/p&gt;
&lt;p&gt;On the Lambda side, &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/aws-lambda-managed-instances?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS added scheduled scaling for functions on Lambda Managed Instances&lt;/a&gt; via EventBridge Scheduler, useful for adjusting capacity ahead of expected traffic, and &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/region-switch-lambda-esm-execution-block?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;ARC Region switch now automates Lambda event source mapping execution during failovers&lt;/a&gt; across Kinesis, DynamoDB Streams, MSK, and SQS, with cross-account support. That last one is very cool. CloudFront got two updates: &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/amazon-cloudfront-mtls-passthrough?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Passthrough Mode for mTLS&lt;/a&gt; that forwards certificates to origins without edge validation, and &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/cloudfront-configurable-premium-flat-rate-plans?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;configurable usage allowances on the Premium flat-rate plan&lt;/a&gt; from 500 million to 6 billion requests and 50 TB to 600 TB per month. And &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/amazon-eventbridge-sdk-integrations?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;EventBridge Scheduler added 619 new SDK API actions across 13 services&lt;/a&gt;, bringing the total coverage to over 270 AWS services.&lt;/p&gt;
&lt;p&gt;Finally, on the security side, &lt;a href=&quot;https://www.wiz.io/blog/introducing-runtime-threat-detection-for-google-cloud-run?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Wiz&#39;s Runtime Sensor for Google Cloud Run is now GA&lt;/a&gt;, with 2000+ detection rules and AI-driven investigation through their Blue Agent. Serverless container monitoring built for serverless containers. Who would have thought?&lt;/p&gt;
&lt;h3&gt;Tutorials&lt;/h3&gt;
&lt;ul&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/database/zero-downtime-dynamodb-construct-migration-from-table-to-tablev2-with-cdk-orphan?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Zero-downtime DynamoDB construct migration: from Table to TableV2 with cdk orphan&lt;/a&gt; by Lee Hannigan&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/database/getting-started-with-change-data-capture-in-amazon-aurora-dsql?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Getting started with Change Data Capture in Amazon Aurora DSQL&lt;/a&gt; by Vijay Karumajji&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://builder.aws.com/content/3DiJq5vQ2hnGq1ddBlCApT8uy6u/dynamic-looping-comes-to-aws-sam?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Dynamic Looping Comes to AWS SAM&lt;/a&gt; by Eric Johnson&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://claude.com/blog/best-practices-for-computer-and-browser-use-with-claude?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Best practices for computer and browser use with Claude&lt;/a&gt; by Lucas Gonzalez and Luca Weihs&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://tricksumo.com/at-most-once-vs-at-least-once-semantics-lambda-durable-function?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AtMostOncePerRetry vs AtLeastOncePerRetry Semantics in Lambda Durable Function Step&lt;/a&gt; by Rishi&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/machine-learning/build-custom-code-based-evaluators-in-amazon-bedrock-agentcore?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Build custom code-based evaluators in Amazon Bedrock AgentCore&lt;/a&gt; by Bharathi Srinivasan&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://serverlessdna.com/strands/ai-assisted-development/layered-configuration-claude-code?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Layered Configuration in Claude Code&lt;/a&gt; by Michael Walmsley&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://monicacolangelo.com/multi-tenant-dynamodb-token-vending-machine?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Per-tenant DynamoDB isolation with the Token Vending Machine pattern&lt;/a&gt; by Monica Colangelo&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://dev.to/aws-builders/lambda-durable-functions-when-you-dont-need-step-functions-20bn?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Lambda Durable Functions, When You Don&#39;t Need Step Functions&lt;/a&gt; by Lewis Sawe&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://medium.com/@RDarrylR/live-canary-deployments-with-aws-sam-the-new-websocket-api-resource-and-lambda-durable-functions-4b029533b34f?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Live Canary Deployments with AWS SAM, the New WebSocket API Resource, and Lambda Durable Functions&lt;/a&gt; by Darryl Ruggles&lt;/li&gt;
&lt;/ul&gt;
&lt;h3&gt;Reads&lt;/h3&gt;
&lt;p&gt;&lt;a href=&quot;https://claude.com/blog/the-founders-playbook?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;The founder&#39;s playbook: Building an AI-native startup&lt;/a&gt;&lt;br /&gt;
Anthropic walks through Idea, MVP, Launch, and Scale for AI-native startups, with real founder stories woven in throughout. Playbooks aren&#39;t the whole answer, but if you&#39;re staring at a blank canvas, this is a pretty good starting point.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://claude.com/blog/how-claude-code-works-in-large-codebases-best-practices-and-where-to-start?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;How Claude Code works in large codebases: Best practices and where to start&lt;/a&gt; by&lt;br /&gt;
A complete walkthrough of Claude Code&#39;s extension points (&lt;a href=&quot;http://claude.md/&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;CLAUDE.md&lt;/a&gt; files, LSP integrations, MCP servers, subagents) and how they shape behavior in enterprise codebases. If you&#39;re getting mediocre results from Claude Code once you push past your toy project, this will probably help you fill some gaps.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://blog.cloudflare.com/cyber-frontier-models?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Project Glasswing: what Mythos showed us&lt;/a&gt; by Grant Bourzikas&lt;br /&gt;
Cloudflare tested Anthropic&#39;s Mythos Preview model on a number of their repos and got firsthand knowledge of why off-the-shelf coding agents fall short. The multi-stage architecture they built is a useful reference for anyone doing serious agentic work outside of the basic coding use case.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://www.nytimes.com/2026/05/18/opinion/ai-boo-commencement-speeches.html?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Opinion | The Generation That Grew Up With A.I. Hates It&lt;/a&gt; by Michelle Goldberg&lt;br /&gt;
Only 18% of Gen Z is hopeful about AI, and 47% of voters under 30 rate it as mostly bad. As someone with two daughters in that demographic, I can&#39;t say I&#39;m surprised. They&#39;ve watched the technology arrive with lots of promises but not a lot of upside for them.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://www.readysetcloud.io/blog/allen.helton/local-agents-scare-me?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Local agents scare me&lt;/a&gt; by Allen Helton&lt;br /&gt;
Allen walks through four attack vectors for local AI agents (shared userland, network adjacency, poisoned context, persistent state) and makes the case that traditional IAM controls don&#39;t fit. Definitely worth reading before you give any agent unrestricted shell access on your machine.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://ranthebuilder.cloud/blog/is-aws-lambda-tenant-isolation-mode-enough-for-saas?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Is AWS Lambda Tenant Isolation Mode Enough for SaaS?&lt;/a&gt; by Ran Isenberg&lt;br /&gt;
Ran breaks down what Lambda&#39;s tenant isolation actually solves and what it doesn&#39;t. The compute side is handled, but data access control is still on you, which has always been the hard part of multi-tenancy.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://medium.com/@siddarthpatil/10-practical-serverless-architecture-lessons-from-aws-summit-london-2026-5ccb29621a27?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;10 Practical Serverless Architecture Lessons from AWS Summit London 2026&lt;/a&gt; by Siddarth Patil&lt;br /&gt;
A grab-bag of serverless patterns from AWS Summit London: Lambda boundaries, async with EventBridge and SQS, cold starts, cost management, and applying the same patterns to GenAI workloads. Most of it is table stakes if you&#39;ve been doing this a while, but the GenAI section is worth a skim.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://architectingautonomy.substack.com/p/cross-domain-governance?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Cross-Domain Governance&lt;/a&gt; by Aaron Sempf&lt;br /&gt;
Aaron explains how autonomous systems should behave when they cross organizational boundaries, proposing monotonic reduction: authority can only be restricted, never amplified, as you move outward. It&#39;s a tidy way to think about a problem most agent platforms haven&#39;t even acknowledged yet. I always feel smarter after reading his stuff.&lt;/p&gt;
&lt;h3&gt;Podcasts, Videos, and more&lt;/h3&gt;
&lt;p&gt;&lt;a href=&quot;https://www.youtube.com/watch?v=mUhYA_Obh-4?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Building Apps with AI + MCP Servers | Serverless Office Hours&lt;/a&gt;&lt;br /&gt;
Brian Zambrano joins Darko Mesaroš to build a serverless application from prompts using Kiro and MCP servers. A solid walkthrough from natural language to deployed AWS infrastructure.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://www.lennysnewsletter.com/p/how-i-ai-html-is-the-new-markdown?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;How I AI: HTML is the new Markdown: How Anthropic engineers are building with Claude Code&lt;/a&gt;&lt;br /&gt;
Claire Vo interviews Thariq Shihipar from Anthropic&#39;s Claude Code team on the shift from Markdown to HTML for AI output, plus patterns like living design systems and micro-apps. I&#39;m all for the HTML move. Markdown files have gotten easier and easier to gloss over, and giving the model a proper display layer makes the output feel a lot less disposable.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://www.youtube.com/watch?v=43tYW89iikU?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Serverless CrAIc Ep84 AI-Generated Code Is a Liability: Technical Debt &amp;amp; Engineering Excellence&lt;/a&gt;&lt;br /&gt;
The Serverless Craic crew digs into the velocity versus debt tradeoff in AI-generated code, including the awkward truth that more tests and more code don’t always mean better quality. The discussion around engineering excellence as a counterweight to AI-driven throughput is where this one gets especially heady. They argue that production code still has to be maintained by humans eventually. Let&#39;s hope.&lt;/p&gt;
&lt;h3&gt;New from AWS&lt;/h3&gt;
&lt;ul&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/aws-cloudformation-cdk-stack?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Reference stack outputs across accounts and Regions with AWS CloudFormation and CDK&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/04/aws-transform-developer-tools?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS Transform agents now available in Kiro, Claude, Cursor, and Codex&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/aws-transform-ai-assistant?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS Transform adds agentic AI assistant to the AWS Toolkit for Visual Studio&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/amazon-ec2-m3-ultra-mac-instances-generally-available?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Announcing general availability of Amazon EC2 M3 Ultra Mac instances&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/amazon-redshift-alter-table-iceberg?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon Redshift adds ALTER TABLE for Iceberg tables and writes via the AWS Glue Data Catalog mount&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/amazon-emr-serverless-aws-regions?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon EMR Serverless is now available in additional AWS Regions&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/aws-announces-AWS-interconnect-multicloud-oci-preview?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS announces AWS Interconnect - multicloud connectivity with Oracle Cloud Infrastructure in preview&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/cloudwatch-logs-query-results?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon CloudWatch Logs announces increased query result limits&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/aws-organizations-increased-scp-quotas?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS Organizations now supports higher quotas for service control policies (SCPs)&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/aws-security-agent-full-repository-code-review?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS Security Agent now supports full repository code reviews&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/amazon-route-53-domains?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon Route 53 Domains adds support for 34 new Top Level Domains including .app, .dev, and .health.&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/aws-sam-cli-cloudformation?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23366&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS SAM CLI adds AWS CloudFormation Language Extensions support to accelerate local serverless development&lt;/a&gt;&lt;/li&gt;
&lt;/ul&gt;
&lt;h3&gt;Final Thoughts 🤔&lt;/h3&gt;
&lt;p&gt;There&#39;s a technical nuance to this Claude Code change that&#39;s worth pulling apart. Running &lt;code&gt;claude -p&lt;/code&gt; isn&#39;t the same as hitting the API directly. Claude Code ships with caching, tool-use optimizations, context management, and prompt structuring that make the interactive product feel as useful as it does. When you wire &lt;code&gt;claude -p&lt;/code&gt; into a script, you get those same optimizations applied to your automated workflows. Hit the raw API yourself and you&#39;re rebuilding all of that from scratch, usually badly, burning extra tokens on every loop, and probably wrecking the economics in the process.&lt;/p&gt;
&lt;p&gt;That&#39;s why this change stings more than a normal pricing adjustment. I understand metering for large-scale programmatic use, but the bigger shift is that now the cheap, optimized path is effectively reserved for interactive use. If you want automation, you either pay API rates or stay inside one of Claude&#39;s tightly controlled (typically not great) interfaces.&lt;/p&gt;
&lt;p&gt;That&#39;s the part that bothers me. You&#39;re paying for more than just access to a raw model. You&#39;re paying for the orchestration layer around it: the caching, context handling, tool execution, prompt shaping, and all the little optimizations that make Claude Code actually useful day to day. Whether those requests originate from a human typing into a terminal or a script running in the background seems mostly immaterial.&lt;/p&gt;
&lt;p&gt;And that&#39;s the bigger question this raises for the industry. Are these systems ultimately meant to become programmable infrastructure, or are they meant to remain interactive products with a human sitting in front of them? Because the economics matter. Automation only works when the cost structure makes sense. If the optimized path is reserved for interactive use while automated use is pushed onto significantly more expensive APIs, then we&#39;re implicitly putting limits on how far these tools can evolve beyond &amp;quot;copilot&amp;quot; workflows.&lt;/p&gt;
&lt;p&gt;That&#39;s worth thinking about before we build entire engineering organizations around them.&lt;/p&gt;
&lt;p&gt;See you next week,&lt;br /&gt;
Jeremy&lt;/p&gt;
&lt;hr /&gt;
&lt;p&gt;I hope you enjoyed this newsletter. We&#39;re always looking for ideas and feedback to make it better and more inclusive, so please feel free to reach out to me via &lt;a href=&quot;https://bsky.app/profile/jeremydaly.com&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Bluesky&lt;/a&gt;, &lt;a href=&quot;http://www.linkedin.com/in/jeremydaly&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;LinkedIn&lt;/a&gt;, &lt;a href=&quot;http://twitter.com/jeremy_daly&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;X&lt;/a&gt;, or &lt;a href=&quot;mailto:contact@jeremydaly.com&quot;&gt;email&lt;/a&gt;.&lt;/p&gt;
</content>
  </entry>
  <entry>
    <title>Issue #365: Valkey 9 Unlocks Hybrid Search on ElastiCache 🔍</title>
    <link href="https://offbynone.io/issues/365/"/>
    <updated>2026-05-12T12:00:00Z</updated>
    <summary>In this issue, Claude Platform sets up shop on AWS, ElastiCache learns to do full-text and hybrid search, and Ampt rolls out Node.js 24 as the default runtime.</summary>
    <id>https://offbynone.io/issues/365/</id>
    <content type="html">&lt;h2&gt;Valkey 9 Unlocks Hybrid Search on ElastiCache 🔍&lt;/h2&gt;
&lt;p&gt;In our &lt;a href=&quot;https://offbynone.io/issues/364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;previous issue&lt;/a&gt;, Amazon Bedrock crossed the final frontier of hosted frontier models, AI agents started buying domain names, and Amazon Q Developer got a one-way ticket to the AWS graveyard. This week, Claude Platform sets up shop on AWS, ElastiCache learns to do full-text and hybrid search, and Ampt rolls out Node.js 24 as the default runtime. Plus, we&#39;ve got plenty of awesome cloud, serverless, and AI content from the community.&lt;/p&gt;
&lt;h3&gt;News &amp;amp; Announcements&lt;/h3&gt;
&lt;p&gt;Anthropic and AWS got even closer this week. &lt;a href=&quot;https://claude.com/blog/claude-platform-on-aws?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23365&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Anthropic introduced the Claude Platform on AWS&lt;/a&gt;, which sits alongside Claude on Bedrock as a second, distinct way to use Claude inside your AWS account. The split is worth understanding: Claude Platform is Anthropic-operated with data processed &lt;em&gt;outside&lt;/em&gt; AWS, while Claude on Bedrock keeps data &lt;em&gt;inside&lt;/em&gt; the AWS boundary. &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/claude-platform-aws?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23365&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Claude Platform on AWS is now generally available&lt;/a&gt; across 18 regions with direct access to Anthropic&#39;s APIs, console, Managed Agents, web search, and prompt caching, all billed through AWS Marketplace. AWS has &lt;a href=&quot;https://aws.amazon.com/blogs/machine-learning/introducing-claude-platform-on-aws-anthropics-native-platform-through-your-aws-account?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23365&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;its own post on the launch&lt;/a&gt; explaining the IAM and Marketplace plumbing. The short version: enterprises that want full Anthropic-native features without leaving their AWS account just got a much cleaner deployment path.&lt;/p&gt;
&lt;p&gt;AgentCore also had a heck of a week. &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/amazon-bedrock-agentcore-runtime?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23365&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AgentCore Runtime now supports bring-your-own file system from S3 and EFS&lt;/a&gt;, letting you mount durable storage directly at agent runtime paths instead of bolting on file access through tools. &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/agentcore-longterm-memory-metadata?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23365&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AgentCore Memory now supports metadata for long-term memory&lt;/a&gt; with up to ten indexed keys that can be set manually or inferred by an LLM, making retrieval over long-term memory actually targetable instead of a vector similarity guessing game. And in the &amp;quot;what could possibly go wrong&amp;quot; category, &lt;a href=&quot;https://aws.amazon.com/blogs/machine-learning/agents-that-transact-introducing-amazon-bedrock-agentcore-payments-built-with-coinbase-and-stripe?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23365&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Bedrock AgentCore Payments launched in preview&lt;/a&gt; (read the &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/04/amazon-bedrock-agentcore-payments-preview?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23365&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;official announcement blog&lt;/a&gt;), built with Coinbase and Stripe and using the x402 protocol to let agents pay for APIs, MCP servers, and web content in stablecoins. So agents now have file systems, memory with metadata, and a wallet. 🔥&lt;/p&gt;
&lt;p&gt;On the agent tooling side, &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/agent-toolkit?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23365&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS announced the Agent Toolkit for AWS&lt;/a&gt;, a managed suite of pre-validated skills for AI coding agents covering application development, data analytics, and AgentCore, with IAM guardrails baked in. Also, &lt;a href=&quot;https://aws.amazon.com/blogs/aws/the-aws-mcp-server-is-now-generally-available?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23365&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;the AWS MCP Server is generally available&lt;/a&gt;, now with IAM context keys, a sandboxed Python execution tool, and better token efficiency. AWS is trying really hard to be the default platform for AI coding agents. Giving devs an opinionated, authenticated entry point seems like the smart play, but AWS doesn&#39;t have the same head start they did with serverless.&lt;/p&gt;
&lt;p&gt;It was a big week for ElastiCache as well. &lt;a href=&quot;https://aws.amazon.com/blogs/database/valkey-turns-two?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23365&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Valkey turned two&lt;/a&gt;, with Docker pulls up 17x year over year and adoption across the major clouds, which is a pretty good trajectory given that it started as a Redis fork barely 24 months ago. They also announced the release of &lt;a href=&quot;https://aws.amazon.com/blogs/database/announcing-valkey-9-0-for-amazon-elasticache?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23365&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Valkey 9.0 for Amazon ElastiCache&lt;/a&gt;, which brings built-in search, hash field expiration, and multi-database support in cluster mode. The headline features got their own announcements: ElastiCache now supports &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/amazon-elasticache-enchanced-search?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23365&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;real-time full-text, exact-match, and numeric range search&lt;/a&gt;, &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/amazon-elasticache-hybrid-search?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23365&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;hybrid search combining vector similarity and full-text&lt;/a&gt;, and &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/amazon-elasticache-aggregations?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23365&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;real-time aggregations&lt;/a&gt;, all at microsecond latency and across all regions at no extra cost. Chaitanya Nuthalapati has a &lt;a href=&quot;https://aws.amazon.com/blogs/database/enhanced-search-for-amazon-elasticache?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23365&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;walkthrough of building search and recommendation engines on top of it&lt;/a&gt; with full code, and there&#39;s a &lt;a href=&quot;https://aws.amazon.com/blogs/database/announcing-aggregations-on-amazon-elasticache?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23365&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;separate post on the aggregations specifically&lt;/a&gt;. ElastiCache is turning into a serious AI workload backend, but it might also be the serverless full-text search service we&#39;ve been waiting for.&lt;/p&gt;
&lt;p&gt;For AWS SAM users, two nice quality-of-life updates: &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/aws-sam-websocket-apis-api-gateway?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23365&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;SAM now natively supports WebSocket APIs for API Gateway&lt;/a&gt;, auto-generating routes, integrations, and IAM permissions from your template, and &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/aws-sam-cli-buildkit-aws-lambda?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23365&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;SAM CLI 1.159.0 added BuildKit support for Lambda container images&lt;/a&gt;, bringing multi-stage builds, better caching, cross-architecture builds, and Docker secrets to the workflow. It seems like these updates should have shipped years ago, but I&#39;m glad to see them land.&lt;/p&gt;
&lt;p&gt;In other Anthropic news, &lt;a href=&quot;https://claude.com/blog/agent-view-in-claude-code?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23365&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Claude Code got agent view&lt;/a&gt;, a centralized UI for managing multiple coding sessions in parallel without juggling terminal tabs. If you&#39;ve been doing this manually with tmux and worktrees, this is going to save you some major pain. Anthropic also rolled out &lt;a href=&quot;https://claude.com/blog/collaborate-with-claude-across-excel-powerpoint-word-and-outlook?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23365&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Claude integrations across Excel, PowerPoint, Word, and Outlook&lt;/a&gt;, with Excel, PowerPoint, and Word now GA and Outlook in public beta. Context follows you across apps, and enterprises get OpenTelemetry logging and Analytics API access for governance. And &lt;a href=&quot;https://claude.com/blog/new-in-claude-managed-agents?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23365&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Claude Managed Agents picked up &amp;quot;dreaming,&amp;quot; outcomes, and multiagent orchestration&lt;/a&gt;, with outcomes being a rubric-based eval system showing up to 10-point improvements on hard tasks. Netflix and Wisedocs are already shipping with it.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://getampt.com/blog/nodejs24-support?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23365&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Ampt now supports Node.js 24&lt;/a&gt; as the default runtime, bringing Web Streams, URLPattern, iterator helpers, and a pile of features that used to require third-party npm packages.&lt;/p&gt;
&lt;p&gt;Finally, &lt;a href=&quot;https://blog.cloudflare.com/building-for-the-future?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23365&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Cloudflare is laying off over 1,100 employees&lt;/a&gt;, which they&#39;re framing as a reorganization for the AI era rather than cost-cutting. The severance package is genuinely good (full base pay through end of 2026 and accelerated equity vesting), but the framing is doing a lot of work. &amp;quot;Reorganization for the AI era&amp;quot; is becoming the corporate euphemism of the decade.&lt;/p&gt;
&lt;h3&gt;Tutorials&lt;/h3&gt;
&lt;ul&gt;
&lt;li&gt;&lt;a href=&quot;https://loige.co/writing-middlewares-for-rust-lambda-functions?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23365&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Writing middlewares for Rust Lambda functions&lt;/a&gt; by Luciano Mammino&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/architecture/choosing-between-single-or-multiple-organizations-in-aws-organizations?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23365&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Choosing between single or multiple organizations in AWS Organizations&lt;/a&gt; by John White&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/database/amazon-aurora-dsql-connections-drivers-strings-and-best-practices?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23365&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon Aurora DSQL connections: Drivers, strings, and best practices&lt;/a&gt; by Rob Petersen&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/database/query-billion-scale-vectors-with-sql-integrating-amazon-s3-vectors-and-aurora-postgresql?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23365&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Query billion-scale vectors with SQL: Integrating Amazon S3 Vectors and Aurora PostgreSQL&lt;/a&gt; by Shayon Sanyal&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://dev.to/robertobelotti/how-i-locked-down-a-static-site-with-lambdaedge-and-cognito-no-backend-required-40el?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23365&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;How I Locked Down a Static Site with Lambda@Edge and Cognito (No Backend Required)&lt;/a&gt; by Roberto Belotti&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/database/migrating-data-from-an-amazon-aurora-snapshot-into-amazon-aurora-dsql?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23365&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Migrating data from an Amazon Aurora snapshot into Amazon Aurora DSQL&lt;/a&gt; by Dan Blaner&lt;/li&gt;
&lt;/ul&gt;
&lt;h3&gt;Reads&lt;/h3&gt;
&lt;p&gt;&lt;a href=&quot;https://chrisebert.net/notes-from-code-with-claude-2026?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23365&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Notes from Code with Claude 2026&lt;/a&gt; by Chris Ebert&lt;br /&gt;
Chris pulls together the announcements that mattered from Code with Claude 2026: the SpaceX compute deal, Multiagent Orchestration, and Dreaming inside Managed Agents. The context window observations are the most useful part for anyone actually shipping agents right now.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://dev.to/thegdsks/aws-lambda-is-dead-the-020-was-never-the-price-2k4j?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23365&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS Lambda Is Dead. The $0.20 Was Never the Price&lt;/a&gt;&lt;br /&gt;
The author migrated 47 Lambda functions to Cloudflare Workers and dropped their monthly bill from $8,362 to $1,790, with most of the savings coming from the orchestration tax (API Gateway, CloudWatch, NAT, egress) rather than Lambda itself. He&#39;s right that the bundle is where the real money goes, and the August 2025 INIT billing change is worth knowing about. But the workloads he&#39;s describing (HTTP APIs, webhooks, auth, edge functions waiting on a database) were never the shape Lambda was built for. Lambda&#39;s actual sweet spot is async event-driven work that needs to fan out to thousands of concurrent executions for seconds at a time, not synchronous request/response paths burning wall clock waiting on Postgres. High-volume systems need to be designed for the runtime you&#39;re putting them on. Putting a sync API behind API Gateway and a NAT&#39;d Lambda and then complaining about the bundle is a design problem dressed up as a pricing problem. Workers is a better fit for that workload, and he should use it. Just don&#39;t declare the tool dead because it was the wrong one for the job.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://www.databricks.com/blog/rethinking-distributed-systems-serverless-performance-and-reliability?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23365&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Rethinking Distributed Systems for Serverless Performance and Reliability&lt;/a&gt; by Aaron Davidson, Roland Fäustlin, and Zach Williams&lt;br /&gt;
Databricks walks through how their serverless Spark platform works, including Spark Connect that decouples apps from clusters, a Serverless Gateway that does the routing, and an autoscaler that earns its name. Using serverless to take 4-5 hour jobs down to 20 minutes is the kind of number that makes the architectural decisions worth reading about.&lt;/p&gt;
&lt;h3&gt;Podcasts, Videos, and more&lt;/h3&gt;
&lt;p&gt;&lt;a href=&quot;https://www.youtube.com/watch?v=irtcYhgQ-vA?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23365&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;How serverless experts build with AI today | Serverless Office Hours&lt;/a&gt;&lt;br /&gt;
Mark Sailes joins Julian Wood to share how serverless experts built Study from Experts, a focused video learning platform for AWS professionals.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://www.youtube.com/watch?v=tKO29SA7CAU?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23365&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Beyond the Basics: Production Serverless Patterns for Extreme Scale • Janak Agarwal • GOTO 2025&lt;/a&gt;&lt;br /&gt;
Janak digs into Lambda patterns that actually hold up under load, with two grounded examples: rapid scale-out for spiky traffic and real-time financial analytics built on Step Functions Distributed Map. This is the kind of content that should be louder than the &amp;quot;Lambda is dead&amp;quot; takes, because it shows what the architecture is genuinely good at.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://www.lennysnewsletter.com/p/spec-driven-development-the-ai-engineering?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23365&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Spec-driven development: The AI engineering workflow at Notion | Ryan Nystrom&lt;/a&gt;&lt;br /&gt;
Claire Vo interviews Ryan Nystrom about how Notion engineers use their internal Boxy system to @mention Codex from comments and get full PRs with screenshots in 20 minutes. The conversation covers practical workflows including configuring subagents, MCP integrations, and the shift toward spec-first development where AI handles implementation.&lt;/p&gt;
&lt;h3&gt;New from AWS&lt;/h3&gt;
&lt;ul&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/aws-waf-dynamic-label-interpolation?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23365&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS WAF introduces dynamic label interpolation for custom request and response handling&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/amazon-aurora-dsql-five-additional-aws-regions?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23365&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon Aurora DSQL is now available in five additional AWS Regions&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/aws-transform-containerization?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23365&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS Transform adds containerization capability during migrations&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/amazon-route-53-resolver-ipv6?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23365&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon Route 53 Resolver endpoints now support additional capabilities for IPv6 query traffic&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/aws-regional-planning-tool-notification?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23365&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS Capabilities by Region now supports availability notifications&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/aws-advanced-jdbc-wrapper-encryption?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23365&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS Advanced JDBC Wrapper now provides client-side encryption&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/concurrencyscaling-support-for-copy?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23365&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon Redshift now scales data ingestion automatically with concurrency scaling for batch workloads&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/amazon-elasticache-cloudwatch-metrics-network-engine-diagnostics?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23365&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon ElastiCache adds thirteen new Amazon CloudWatch metrics for network capacity planning and engine diagnostics&lt;/a&gt;&lt;/li&gt;
&lt;/ul&gt;
&lt;h3&gt;Final Thoughts 🤔&lt;/h3&gt;
&lt;p&gt;Look at what AWS shipped this week and squint a little. Claude Platform on AWS, Agent Toolkit, and AWS MCP Server GA, plus AgentCore gets durable file systems, metadata for long-term memory, and payments with stablecoin rails. AWS is staking out the substrate layer for the agentic era, and the feature list isn&#39;t random.&lt;/p&gt;
&lt;p&gt;The bet is straightforward. If agents need compute, identity, storage, memory, payment, and an authenticated way to call services, AWS already has four of those and is shipping the other two as fast as they can write their press releases. The pitch to enterprises is: your agents already run on AWS, your data already lives on AWS, your IAM already governs everything, so why would you run the agent loop anywhere else?&lt;/p&gt;
&lt;p&gt;It&#39;s a credible play. But the serverless comparison I mentioned earlier is the one worth thinking about. AWS had a multi-year head start with Lambda, and the platform shape was so unfamiliar that competitors took years to even define the category. Agents don&#39;t have that property. Cloudflare, Vercel, Modal, Fly, and a dozen smaller platforms are already shipping agent primitives. The Anthropic-AWS deal is notable, but Anthropic will sell its service to anyone willing to buy. Model providers are commodity inputs now. The differentiation has to come from somewhere else.&lt;/p&gt;
&lt;p&gt;The substrate fight will be won on governance, observability, and cost controls, not raw capability. Every platform is going to give agents file systems and wallets and OS-level actions. The platform that wins is the one where, when an agent does something dumb or expensive at 3 a.m., you can see exactly what happened, who authorized it, what it cost, and how to stop it from happening again. AWS has decades of muscle memory on that exact problem, which is their edge.&lt;/p&gt;
&lt;p&gt;If you&#39;re building on any of these primitives, the planning question is no longer &amp;quot;can the agent do this.&amp;quot; It&#39;s &amp;quot;when this agent does something I didn&#39;t expect, what&#39;s my blast radius and how fast can I close it.&amp;quot; Build for that and the rest takes care of itself.&lt;/p&gt;
&lt;p&gt;See you next week,&lt;br /&gt;
Jeremy&lt;/p&gt;
&lt;hr /&gt;
&lt;p&gt;I hope you enjoyed this newsletter. We&#39;re always looking for ideas and feedback to make it better and more inclusive, so please feel free to reach out to me via &lt;a href=&quot;https://bsky.app/profile/jeremydaly.com&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Bluesky&lt;/a&gt;, &lt;a href=&quot;http://www.linkedin.com/in/jeremydaly&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;LinkedIn&lt;/a&gt;, &lt;a href=&quot;http://twitter.com/jeremy_daly&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;X&lt;/a&gt;, or &lt;a href=&quot;mailto:contact@jeremydaly.com&quot;&gt;email&lt;/a&gt;.&lt;/p&gt;
</content>
  </entry>
  <entry>
    <title>Issue #364: Agents With Credit Cards 🛒</title>
    <link href="https://offbynone.io/issues/364/"/>
    <updated>2026-05-05T12:00:00Z</updated>
    <summary>In this issue, Amazon Bedrock crosses the final frontier of hosted frontier models, AI agents can now buy domain names for side projects they&#39;ll never finish, and Amazon Q Developer gets a one-way ticket to the AWS graveyard.</summary>
    <id>https://offbynone.io/issues/364/</id>
    <content type="html">&lt;h2&gt;Agents With Credit Cards 🛒&lt;/h2&gt;
&lt;p&gt;In our &lt;a href=&quot;https://offbynone.io/issues/363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;previous issue&lt;/a&gt;, serverless became less stateless, OpenAI dropped two major model upgrades, and Claude went after creatives. This week, Amazon Bedrock crosses the final frontier of hosted frontier models, AI agents can now buy domain names for side projects they&#39;ll never finish, and Amazon Q Developer gets a one-way ticket to the AWS graveyard. Plus, we&#39;ve got lots of amazing cloud, serverless, and AI content from the community.&lt;/p&gt;
&lt;h3&gt;News &amp;amp; Announcements&lt;/h3&gt;
&lt;p&gt;In AWS news, &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/aurora-dsql-json-support?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon Aurora DSQL now supports the JSON data type with compression&lt;/a&gt;, which is a great addition that pushes DSQL closer to Postgres-style storage semantics. Over on the edge, &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/amazon-cloudfront-websockets-vpc-origins?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon CloudFront announced WebSocket support for VPC origins&lt;/a&gt;, letting you keep origins secured inside the VPC while still allowing WebSocket traffic through. CloudFront also &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/04/cloudfront-invalidation-cache-tag?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;now supports invalidation by cache tag&lt;/a&gt;, which is a really big win. If you wanted to invalidate groups of files before, you had to specify all the URL patterns yourself and keep track of them. Tag-based invalidation lets you flush a logical batch of files without nuking the entire cache, which is way cheaper and more efficient.&lt;/p&gt;
&lt;p&gt;The agent autonomy story keeps getting bigger (and scarier). AWS announced that &lt;a href=&quot;https://aws.amazon.com/blogs/aws/modernize-your-workflows-amazon-workspaces-now-gives-ai-agents-their-own-desktop-preview?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon WorkSpaces now gives AI agents their own desktop in preview&lt;/a&gt;. If you still have your inventory managed with Microsoft Access on Windows 95, then this might be for you. We&#39;re slowly starting to treat AI agents as independent, autonomous things with increasingly more permissive sandboxes. That has real upside, but also real downside risk. Pair that with &lt;a href=&quot;https://aws.amazon.com/blogs/machine-learning/introducing-os-level-actions-in-amazon-bedrock-agentcore-browser?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;OS Level Actions in Amazon Bedrock AgentCore Browser&lt;/a&gt;, which lets agents interact with native popups and dialogs that previously blocked browser automation, and the sandbox metaphor gets thinner every minute. Cloudflare is on the same trajectory: &lt;a href=&quot;https://blog.cloudflare.com/agents-stripe-projects?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;agents can now create Cloudflare accounts, buy domains, and deploy&lt;/a&gt;, which is impressive, but means an agent that can stand up infrastructure is also an agent that can run up your cloud bills.&lt;/p&gt;
&lt;p&gt;Inside Bedrock AgentCore itself there was a steady stream of updates. &lt;a href=&quot;https://aws.amazon.com/blogs/machine-learning/introducing-the-agent-quality-loop-agentcore-optimization-now-in-preview?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AgentCore Optimization is now in preview&lt;/a&gt;, allowing agents to improve production performance by analyzing their own traces. &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/04/amazon-bedrock-agentcore?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AgentCore Identity now supports On-Behalf-Of token exchange&lt;/a&gt;, letting an agent log in as a delegated human user, which is again powerful and a little terrifying. And &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/04/amazon-bedrock-agentcore-runtime?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AgentCore Runtime now supports Node.js for direct code deployment&lt;/a&gt;, so you can ship Node agents as ZIP uploads with bundled &lt;code&gt;node_modules&lt;/code&gt; instead of needing a container. Also, &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/04/bedrock-openai-models-codex-managed-agents?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Bedrock now offers OpenAI models, Codex, and Managed Agents in limited preview&lt;/a&gt;, which means Bedrock now hosts effectively every major frontier model.&lt;/p&gt;
&lt;p&gt;On the compute and tooling side, &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/04/aws-lambda-adds-ruby?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS Lambda added support for Ruby 4.0&lt;/a&gt;. AWS is also leaning hard into Amazon Quick. You can now &lt;a href=&quot;https://aws.amazon.com/blogs/machine-learning/generate-dashboards-from-natural-language-prompts-in-amazon-quick?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;generate dashboards from natural language prompts&lt;/a&gt; and &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/04/amazon-quick-macos-windows-preview?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;it&#39;s now available as a desktop application for macOS and Windows in preview&lt;/a&gt;. Meanwhile, &lt;a href=&quot;https://aws.amazon.com/blogs/devops/amazon-q-developer-end-of-support-announcement?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon Q Developer got an end-of-support announcement&lt;/a&gt;, which we all knew was coming. Q Developer was a waypoint along AWS&#39;s agentic coding journey, not the destination. And the &lt;a href=&quot;https://aws.amazon.com/blogs/compute/serverless-icymi-q1-2026?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Serverless ICYMI Q1 2026 roundup&lt;/a&gt; is worth a look. Lots of interesting stuff including durable function updates, larger Lambda, SQS, and EventBridge payloads, DynamoDB cross-account replication, and a bunch of AgentCore infrastructure work.&lt;/p&gt;
&lt;p&gt;In Anthropic news, &lt;a href=&quot;https://claude.com/blog/claude-security-public-beta?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Claude Security is now in public beta&lt;/a&gt;, which scans codebases for vulnerabilities by inspecting how components interact rather than pattern-matching against a CVE list. They&#39;ve already tested it with hundreds of organizations over the past two months, and the approach is impressive. Also, the &lt;a href=&quot;https://claude.com/blog/claude-api-skill?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Claude API skill is now available in CodeRabbit, JetBrains, Resolve AI, and Warp&lt;/a&gt;, bundling production-ready knowledge of API patterns, prompt caching rules, and per-model configuration directly into those tools and staying current as you work.&lt;/p&gt;
&lt;p&gt;Finally, on the Cloudflare side, they &lt;a href=&quot;https://blog.cloudflare.com/dynamic-workflows?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;introduced Dynamic Workflows&lt;/a&gt;, which combines durable execution with dynamic Workers so the platform can route workflow instances to different tenant code without pre-deployed targets. It&#39;s another interesting AI-agent primitive, especially for things like per-tenant CI/CD pipelines.&lt;/p&gt;
&lt;h3&gt;Tutorials&lt;/h3&gt;
&lt;ul&gt;
&lt;li&gt;&lt;a href=&quot;https://theburningmonk.com/2026/05/inbox-outbox-patterns-for-reliable-event-processing?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Inbox &amp;amp; Outbox patterns for reliable event processing&lt;/a&gt; by Yan Cui&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/machine-learning/organizing-agents-memory-at-scale-namespace-design-patterns-in-agentcore-memory?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Organizing Agents’ memory at scale: Namespace design patterns in AgentCore Memory&lt;/a&gt; by Noor Randhawa&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/machine-learning/run-custom-mcp-proxies-serverless-on-amazon-bedrock-agentcore-runtime?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Run custom MCP proxies serverless on Amazon Bedrock AgentCore Runtime&lt;/a&gt; by Nizar Kheir&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://www.sls.guru/blog/before-you-rebuild-your-rag-stack-7-reasons-your-answers-are-weak-its-not-the-model?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Before You Rebuild Your RAG Stack: Why Your Answers are Weak | Serverless Guru&lt;/a&gt; by Cyril Bandolo&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://darryl-ruggles.cloud/s3-files-the-end-of-download-process-upload-with-terraform?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;S3 Files: Simplified AWS Lambda Processing with Terraform&lt;/a&gt; by Darryl Ruggles&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://dev.to/accreditly/replacing-puppeteer-on-aws-lambda-for-screenshots-3622?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Replacing Puppeteer on AWS Lambda for Screenshots&lt;/a&gt; by Mike Griffiths&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://dev.to/aws-builders/how-i-used-amazon-quick-to-run-a-full-security-audit-on-my-saas-and-fixed-11-vulnerabilities-in-4n8o?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;How I Used Amazon Quick to Run a Full Security Audit on My SaaS — and Fixed 11 Vulnerabilities in One Session&lt;/a&gt; by Asad Marcus&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://dev.to/aws-builders/i-injected-three-faults-the-agent-found-all-of-them-5pi?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;I Injected Three Faults. The Agent Found All of Them.&lt;/a&gt; by Romar Cablao&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/database/building-agentic-ai-for-amazon-rds-for-sql-server-with-strands-and-agentcore?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Building agentic AI for Amazon RDS for SQL Server with Strands and AgentCore&lt;/a&gt; by Sudhir Amin&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://dev.to/rdarrylr/its-all-about-that-memory-using-long-and-short-term-memory-with-agents-2m21?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;It&#39;s All About That Memory - Using Long and Short Term Memory with Agents&lt;/a&gt; by Darryl Ruggles&lt;/li&gt;
&lt;/ul&gt;
&lt;h3&gt;Reads&lt;/h3&gt;
&lt;p&gt;&lt;a href=&quot;https://claude.com/blog/lessons-from-building-claude-code-prompt-caching-is-everything?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Lessons from building Claude Code: Prompt caching is everything&lt;/a&gt;&lt;br /&gt;
The Claude Code team treats prompt cache hit rate as an SRE metric with SEV alerts, because caching&#39;s prefix-match rule makes obvious optimizations backfire: switching to Haiku mid-session for an easy question costs more than letting Opus answer it. The post covers the patterns that follow, including modeling Plan Mode as tools, deferring MCP schemas via stubs, and cache-safe forking for compaction.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://architectingautonomy.substack.com/p/the-reinvention-problem?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;The Reinvention Problem&lt;/a&gt;&lt;br /&gt;
Hans Schabert and Aaron Sempf ran the same prescribed agent procedure hundreds of times and watched it splinter into dozens of execution paths, with the most common one accounting for barely a quarter of runs. Their argument: stuffing a workflow into a system prompt hands the model a reference manual when what governance actually requires is an order, and no amount of better prompting or larger context will close that gap.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://heeki.medium.com/interrupting-agents-with-human-in-the-loop-feedback-c46e806d36fe?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Interrupting agents with human-in-the-loop feedback&lt;/a&gt;&lt;br /&gt;
Heeki Park catalogs four ways to wedge human approval into an agent before it issues a refund or revokes access: model-moderated inline functions in AgentCore harness, Strands BeforeToolCallEvent hooks, in-tool ctx.interrupt() calls, and MCP server elicitations. Each comes with code samples and a clear &amp;quot;when to use&amp;quot; rubric depending on whether tool names are known upfront and who owns the tool code.&lt;/p&gt;
&lt;h3&gt;Podcasts, Videos, and more&lt;/h3&gt;
&lt;p&gt;&lt;a href=&quot;https://www.youtube.com/watch?v=FD1iDCj2r5Q?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Automating AWS Lambda runtime upgrades | Serverless Office Hours&lt;/a&gt;&lt;br /&gt;
Dan Fox and Brian Krygsman join Julian Wood to explore how AWS Transform custom can take the pain out of Lambda runtime migrations. They cover AWS Transform custom, a tool for automating Lambda runtime upgrades, and walk through how the AI agent manages code changes, dependency updates, and validation when migrating from deprecated to modern runtimes.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://www.youtube.com/watch?v=y4517SxUH_s?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Serverless &amp;amp; OpenTelemetry ❤️ Better Together&lt;/a&gt;&lt;br /&gt;
James Eastham shows you how to escape the pain of clicking through endless CloudWatch log groups and trying to piece together X-Ray by learning how to instrument your .NET serverless apps with OpenTelemetry.&lt;/p&gt;
&lt;h3&gt;New from AWS&lt;/h3&gt;
&lt;ul&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/workspaces-ai-agents?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon WorkSpaces now lets AI agents operate desktop applications (Preview)&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/aws-iam-increased-quotas?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS IAM now provides higher maximum quotas for roles, role trust policies, instance profiles, managed policies, and identity providers&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/amazon-eventbridge-data-aws-cloudtrail?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon EventBridge supports data plane logging to AWS CloudTrail&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/05/amazon-cloudwatch-logs-query-by-tags?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon CloudWatch Logs Insights supports querying by log group tags&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/04/amazon-opensearch-service-supports-index-level-encryption?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon OpenSearch Service now supports index-level encryption&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/04/amazon-cloudwatch-agent-ec2?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon CloudWatch adds visual agent configuration to the EC2 console&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/04/amazon-ecs-mi-gpu-metrics?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon ECS Managed Instances now supports NVIDIA GPU metrics&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/04/paraphrase-multilingual-table-transformer-bielik-on-sagemaker-jumpstart?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Paraphrase-multilingual-MiniLM-L12-v2, Table Transformer Detection, and Bielik-11B-v3.0-Instruct are now available in Amazon SageMaker JumpStart&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/04/gemma-4-models-on-sagemaker-jumpstart?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Gemma 4 models are now available in Amazon SageMaker JumpStart&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/04/quick-sharepoint-access-control?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon Quick now supports document-level access controls for SharePoint knowledge bases&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/04/custom-applications?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Build custom applications using natural language in Amazon Quick (Preview)&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/04/amazon-quick-google-workspace-zoom?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon Quick expands integrations to include Google Workspace, Zoom, Airtable, and more&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/04/amazon-quick-free-plus?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Start using Amazon Quick for free in minutes with Free and Plus pricing plans&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/04/amazon-quick?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23364&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon Quick now supports document and visual creation in chat&lt;/a&gt;&lt;/li&gt;
&lt;/ul&gt;
&lt;h3&gt;Final Thoughts 🤔&lt;/h3&gt;
&lt;p&gt;Agents now get their own Windows desktops. They can buy domains, spin up Cloudflare accounts, deploy infrastructure, dismiss native OS dialogs, and impersonate users via delegated tokens. A year ago we were arguing about whether agents should be allowed to run shell commands. Now AWS is handing them WorkSpaces and Cloudflare is handing them credit cards. The sandbox keeps getting roomier, and the blast radius keeps growing with it.&lt;/p&gt;
&lt;p&gt;I&#39;m not against any of this. The capability story is genuinely exciting, and most of these primitives are things real production systems need. But we&#39;re shipping the autonomy faster than the controls. On-Behalf-Of token exchange in AgentCore Identity is a great example: powerful for legitimate delegation, also a fantastic way to lose the audit trail if you&#39;re not careful about how you scope it. Same story with agents that can stand up cloud accounts. Great until one of them runs a runaway loop on your billing.&lt;/p&gt;
&lt;p&gt;The Bedrock news is the other shoe dropping. Adding OpenAI models, Codex, and Managed Agents in preview means Bedrock is now the universal hosting layer for frontier models. That&#39;s a real shift. Model choice is becoming an AWS configuration setting rather than a vendor commitment, which is good for builders and very interesting for the rest of the market.&lt;/p&gt;
&lt;p&gt;The pattern across all of this is clear: the platforms are racing to give agents more rope, and the governance, observability, and cost-control story is still catching up. If you&#39;re building on these primitives, that gap is where you live now. Plan for it.&lt;/p&gt;
&lt;p&gt;See you next week,&lt;br /&gt;
Jeremy&lt;/p&gt;
&lt;hr /&gt;
&lt;p&gt;I hope you enjoyed this newsletter. We&#39;re always looking for ideas and feedback to make it better and more inclusive, so please feel free to reach out to me via &lt;a href=&quot;https://bsky.app/profile/jeremydaly.com&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Bluesky&lt;/a&gt;, &lt;a href=&quot;http://www.linkedin.com/in/jeremydaly&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;LinkedIn&lt;/a&gt;, &lt;a href=&quot;http://twitter.com/jeremy_daly&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;X&lt;/a&gt;, or &lt;a href=&quot;mailto:contact@jeremydaly.com&quot;&gt;email&lt;/a&gt;.&lt;/p&gt;
</content>
  </entry>
  <entry>
    <title>Issue #363: Serverless Isn&#39;t Stateless Anymore 💾</title>
    <link href="https://offbynone.io/issues/363/"/>
    <updated>2026-04-28T12:00:00Z</updated>
    <summary>In this issue, serverless becomes less stateless, OpenAI drops two major model upgrades, and Claude goes after creatives.</summary>
    <id>https://offbynone.io/issues/363/</id>
    <content type="html">&lt;h2&gt;Serverless Isn&#39;t Stateless Anymore 💾&lt;/h2&gt;
&lt;p&gt;In our &lt;a href=&quot;https://offbynone.io/issues/362&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;previous issue&lt;/a&gt;, Claude got a major upgrade, AWS made AI costs more visible, and Cloudflare went all-in on agents. This week, serverless becomes less stateless, OpenAI drops two major model upgrades, and Claude goes after creatives. Plus, we&#39;ve got plenty of content from the cloud, serverless, and AI communities.&lt;/p&gt;
&lt;h3&gt;News &amp;amp; Announcements&lt;/h3&gt;
&lt;p&gt;Maybe you noticed that AWS is turning serverless into something a lot more… stateful. &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/04/aws-lambda-amazon-s3?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Lambda can now mount S3 as a file system with S3 Files&lt;/a&gt;, which is a pretty big shift in how you think about data access in functions. Pair that with the &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/04/lambda-durable-execution-java-ga?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Lambda Durable Execution SDK for Java going GA&lt;/a&gt; and &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/04/lambda-durable-functions-16-new-regions?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;durable functions expanding to 16 more regions&lt;/a&gt;, and it’s clear AWS is moving Lambda toward long-running, stateful workflows without giving up the &amp;quot;serverless&amp;quot; model.&lt;/p&gt;
&lt;p&gt;On the agent side, AWS continues its work to remove developer friction. The &lt;a href=&quot;https://aws.amazon.com/blogs/machine-learning/get-to-your-first-working-agent-in-minutes-announcing-new-features-in-amazon-bedrock-agentcore?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;latest Amazon Bedrock AgentCore updates&lt;/a&gt; promise you can get a working agent running in minutes, with new capabilities around orchestration, tooling, and faster setup. That’s backed by additional &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/04/agentcore-new-features-to-build-agents-faster?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AgentCore feature releases&lt;/a&gt; and infrastructure improvements like &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2024/04/agentcore-gateway-identity-vpc?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Gateway + Identity support for VPC egress&lt;/a&gt;, which handles one of the more annoying real-world constraints when connecting agents to private systems.&lt;/p&gt;
&lt;p&gt;AWS and Anthropic also continue to get closer. There’s an &lt;a href=&quot;https://www.anthropic.com/news/anthropic-amazon-compute?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;expanded partnership for massive new compute capacity&lt;/a&gt;, and you can now run &lt;a href=&quot;https://aws.amazon.com/blogs/machine-learning/from-developer-desks-to-the-whole-organization-running-claude-cowork-in-amazon-bedrock?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Claude Cowork directly in Amazon Bedrock&lt;/a&gt;. I still think this is a great bet by AWS to own the integration point for the AI model ecosystem.&lt;/p&gt;
&lt;p&gt;After last week&#39;s Opus 4.7 announcement, you knew it wouldn&#39;t be long before OpenAI responded. &lt;a href=&quot;https://openai.com/index/introducing-gpt-5-5?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;GPT-5.5&lt;/a&gt; is here with all the expected benchmark wins and a 1M token context window, which is starting to feel less like a flex and more like table stakes. They also dropped &lt;a href=&quot;https://openai.com/index/introducing-chatgpt-images-2-0?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;ChatGPT Images 2.0&lt;/a&gt;, which is scary good. Alongside that, we got &lt;a href=&quot;https://openai.com/index/introducing-workspace-agents-in-chatgpt?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;workspace agents in ChatGPT&lt;/a&gt;, more signs of the &lt;a href=&quot;https://openai.com/index/next-phase-of-microsoft-partnership?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;next phase of the Microsoft partnership&lt;/a&gt;, and a fresh set of &lt;a href=&quot;https://openai.com/index/our-principles?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;“principles”&lt;/a&gt; to remind us everything is under control. 😳&lt;/p&gt;
&lt;p&gt;Anthropic isn&#39;t slowing down either. They just announced &lt;a href=&quot;https://www.anthropic.com/news/claude-for-creative-work?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Claude for Creative Work&lt;/a&gt;, which includes new plugins and integrations with partners like Blender, Autodesk, Adobe, Ableton, and Splice. These are tools that let Claude work directly alongside the software creative professionals are using every day. Their strategy is absolutely 🔥. They’re also rolling out &lt;a href=&quot;https://claude.com/blog/claude-managed-agents-memory?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;built-in memory for Claude managed agents&lt;/a&gt;, now in public beta. Memory is quickly becoming the differentiator, and everyone is racing to make it feel less like a hack and more like infrastructure.&lt;/p&gt;
&lt;h3&gt;Tutorials&lt;/h3&gt;
&lt;ul&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/database/dsql-sql-dialect-how-amazon-aurora-dsql-differs-from-single-instance-postgresql?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;DSQL SQL Dialect: How Amazon Aurora DSQL differs from single-instance PostgreSQL&lt;/a&gt; by Rob Petersen&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://dev.to/tanseer/your-aws-cognito-emails-are-going-to-spam-here-is-how-to-fix-it-step-by-step-4989?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Your AWS Cognito Emails Are Going to Spam — Here Is How to Fix It Step by Step&lt;/a&gt; by Tanseer&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://dev.to/abhishek_gupta_pinpo/dynamodb-vs-rds-at-10k-100k-and-1m-rps-a-pre-deployment-simulation-comparison-3eco?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;DynamoDB vs RDS at 10K, 100K, and 1M RPS: a pre-deployment simulation comparison&lt;/a&gt; by Abhishek Gupta&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://dev.to/aws-heroes/serverless-applications-on-aws-with-lambda-using-java-25-api-gateway-and-aurora-dsql-part-6-34ni?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Serverless applications on AWS with Lambda using Java 25, API Gateway and Aurora DSQL - Part 6 Using GraalVM Native Image&lt;/a&gt; by Vadym Kazulkin&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/database/best-practices-and-architecture-patterns-for-cross-account-sharing-in-oracle-databaseaws?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Best practices and architecture patterns for cross-account sharing in Oracle Database@AWS&lt;/a&gt; by Yamuna Palasamudram&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/machine-learning/cost-effective-multilingual-audio-transcription-at-scale-with-parakeet-tdt-and-aws-batch?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Cost-effective multilingual audio transcription at scale with Parakeet-TDT and AWS Batch&lt;/a&gt; by Gleb Geinke&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/machine-learning/build-strands-agents-with-sagemaker-ai-models-and-mlflow?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Build Strands Agents with SageMaker AI models and MLflow&lt;/a&gt; by Dheeraj Hegde&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://studyfromexperts.com/blogs/securing-private-video-content-with-cloudfront-signed-urls-and-serverless-on-aws?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Securing Private Video Content with CloudFront Signed URLs and Serverless on AWS&lt;/a&gt; by Lee Gilmore&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://heeki.medium.com/building-an-agent-harness-31942331d605?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Building an agent harness&lt;/a&gt; by Heeki Park&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://developers.openai.com/api/docs/guides/deployment-checklist?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;OpenAI API deployment checklist&lt;/a&gt; by OpenAI&lt;/li&gt;
&lt;/ul&gt;
&lt;h3&gt;Reads&lt;/h3&gt;
&lt;p&gt;&lt;a href=&quot;https://www.linkedin.com/pulse/anthropic-opus-46-vs-47-which-better-code-quality-experiment-goh-uroxc?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Anthropic Opus 4.6 vs 4.7 - Which is better? A code quality experiment.&lt;/a&gt;&lt;br /&gt;
An AWS AI Hero tests Claude Opus 4.6 against 4.7 using the same Tetris implementation requirements across 13 code quality dimensions. Some folks are calling 4.6 a step back, but 4.7 seems to be finding its footing. I’ve been pretty happy with it so far. Feels like a reminder that model progress isn’t always a straight line, but the trajectory still points up.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://dev.to/riya_mittal_cdd264250ad45/serverless-finops-why-lambda-cost-models-break-every-assumption-you-learned-from-vms-42c5?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Serverless FinOps: Why Lambda Cost Models Break Every Assumption You Learned from VMs&lt;/a&gt;&lt;br /&gt;
Riya Mittal explains how Lambda&#39;s three-dimensional pricing (invocations, duration, memory) creates a fundamentally different cost model than VMs. Keeping cost top of mind is table stakes now. But optimizing for cost alone misses the bigger picture. Scale, performance, and operational overhead all show up eventually. The real game is balancing all three without painting yourself into a corner.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://claude.com/blog/building-agents-that-reach-production-systems-with-mcp?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Building agents that reach production systems with MCP&lt;/a&gt;&lt;br /&gt;
Nice breakdown of three different ways to wire systems into MCP servers. More importantly, it’s another example of patterns starting to solidify. Still early, still messy, but the industry is slowly converging on what “good” looks like.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://edjgeek.com/blog/lambda-cold-starts-dead?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Cold Starts Are Dead&lt;/a&gt;&lt;br /&gt;
Cold starts aren’t what they used to be. There are still edge cases, but for most workloads, they’re manageable or negligible. Eric Johnson covers how platform improvements and better patterns minimize them, resulting in them rarely showing up where it actually matters.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://openai.com/index/speeding-up-agentic-workflows-with-websockets?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Speeding up agentic workflows with WebSockets in the Responses API&lt;/a&gt;&lt;br /&gt;
OpenAI explains how they reduced agentic workflow latency by 40% using WebSockets instead of repeated HTTP requests. The technical approach maintains persistent connections and caches conversation state, eliminating redundant processing of conversation history while exposing the full speed of their faster GPT-5.3-Codex-Spark model.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://teriradichel.substack.com/p/reducing-token-burn-rate-with-a-well?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Reducing Token Burn Rate With A Well-Designed Architecture&lt;/a&gt;&lt;br /&gt;
Teri Radichel walks through building a Lambda troubleshooting system that separates deterministic data gathering from AI analysis. The approach avoids burning tokens on repetitive queries by using traditional code to collect logs and configuration, only invoking AI for interpretation. Stop wasting tokens on repetitive work and only pay for actual insight.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://tylerfolkman.substack.com/p/i-run-qwen-36-on-two-gpus-because?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;I Run Qwen 3.6 on Two GPUs Because Renting AI Is Boring&lt;/a&gt;&lt;br /&gt;
Tyler Folkman explains that locally hosted models might not match the top-tier APIs, but they have one big advantage. They don’t go down. While Anthropic and others keep having “moments” (like as I&#39;m writing this), running your own stack looks a lot less boring.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://architectingautonomy.substack.com/p/the-escalation-trap?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;The Escalation Trap&lt;/a&gt;&lt;br /&gt;
Aaron Sempf walks through three failure modes of human escalation in AI systems: over-escalation creating bottlenecks, selective escalation missing new edge cases, and avoiding escalation entirely. The piece argues for moving escalation decisions to a separate governance layer that evaluates authority boundaries before execution. This is a hard problem.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://ranthebuilder.cloud/blog/how-i-use-claude-cowork-to-write-with-ai-in-my-voice?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;How I Use Claude Cowork To Write With AI In My Voice&lt;/a&gt;&lt;br /&gt;
Ran Isenberg walks through his Claude Cowork configuration for generating content that sounds human. But trying too hard to “not sound like AI” can backfire. A lot of the things people avoid, like short sentences and clarity, are just good writing. The goal shouldn&#39;t be to hide AI; it should be to help you articulate &lt;em&gt;your thoughts and ideas&lt;/em&gt; clearly.&lt;/p&gt;
&lt;h3&gt;Podcasts, Videos, and more&lt;/h3&gt;
&lt;p&gt;&lt;a href=&quot;https://www.youtube.com/watch?v=8rjKqb79Qyg?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Serverless CrAIc Ep 83 Psychological Safety in the AI Era (No One Talks About This)&lt;/a&gt;&lt;br /&gt;
Serverless CrAIc explores how rapid AI adoption challenges team dynamics, mentorship capacity, and organizational culture. Keeping up used to be hard, but now it’s relentless. Fast doesn’t guarantee success, but it does help with learning. And that’s the frustrating part. Watching others move quickly and wondering what they’ve figured out that you haven’t. Also, no, we’re probably not all losing our jobs tomorrow. But it’s not crazy to wonder what the people in charge think.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://www.youtube.com/watch?v=KRT0Z7k01GE?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS Lambda durable functions: Best Practices, AI patterns, and Futures | Serverless Office Hours&lt;/a&gt;&lt;br /&gt;
Michael Gasch and Eric Johnson join Julian Wood to explore the latest in AWS Lambda durable functions, from Java SDK GA, S3 File support, to what&#39;s coming next.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://www.lennysnewsletter.com/p/how-anthropics-product-team-moves?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;How Anthropic’s product team moves faster than anyone else | Cat Wu (Head of Product, Claude Code)&lt;/a&gt;&lt;br /&gt;
Lenny&#39;s interview with Cat Wu explores how Anthropic builds products in days rather than months, and why their employees build custom internal tools instead of buying SaaS. Lots of great insights in here, including emerging PM skills in AI and the shift toward managing AI agent fleets rather than doing tasks yourself.&lt;/p&gt;
&lt;h3&gt;New from AWS&lt;/h3&gt;
&lt;ul&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/04/amazon-msk-serverless-13-regions?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon MSK Serverless expands to 13 new AWS regions&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/04/qwen-models-on-sagemaker-jumpstart?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Five new Qwen models for coding agents and efficient reasoning are now available in Amazon SageMaker JumpStart&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/03/aws-data-exports-cross-account-delivery-cost?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS Data Exports now supports cross-account delivery&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/04/backup-policies-aurora-dsql-redshift-serverless?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS Backup adds Amazon Redshift Serverless and Aurora DSQL support for AWS Organizations backup policies&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/04/cloudwatch-logs-insights-join-sub-query?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon CloudWatch Logs Insights introduces JOIN and sub-query commands&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/04/amazon-location-service-bulk-address-validation?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon Location Service now offers bulk address validation for the United States, Canada, Australia, and the United Kingdom&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/04/s3-five-additional-checksum-algorithms?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon S3 now supports five additional checksum algorithms&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/04/amazon-redshift-serverless-ai-driven-scaling-default?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon Redshift Serverless AI-driven scaling is now the default for new workgroups&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/04/redshift-update-delete-merge-iceberg-tables?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon Redshift supports UPDATE, DELETE, MERGE for Apache Iceberg tables&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/04/amazon-sagemaker-ft-qwen3-5?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon SageMaker AI now supports serverless model customization for Qwen3.5 models&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/04/amazon-smus-ci-cd-cli?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon SageMaker Unified Studio now offers CI/CD CLI for data and AI applications&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/04/sagemaker-ai-inference-rec?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon SageMaker AI launches optimized generative AI inference recommendations&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/04/amazon-athena?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon Athena simplifies federated queries with managed connectors&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/04/aws-compute-optimizer-ec2-rds?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS Compute Optimizer supports 162 new EC2 instance types and 32 new RDS DB instance classes&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/machine-learning/amazon-sagemaker-ai-now-supports-optimized-generative-ai-inference-recommendations?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon SageMaker AI now supports optimized generative AI inference recommendations&lt;/a&gt;&lt;/li&gt;
&lt;/ul&gt;
&lt;h3&gt;Developer Tools&lt;/h3&gt;
&lt;p&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/developer/create-minimal-reproductions-for-aws-sdk-javascript-v3-with-create-aws-sdk-repro?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23363&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Create minimal reproductions for AWS SDK JavaScript v3 with create-aws-sdk-repro&lt;/a&gt; by John Lwin&lt;br /&gt;
AWS released create-aws-sdk-repro, a CLI tool that generates boilerplate for AWS SDK for JavaScript v3 projects. It handles service selection, environment setup (Node.js, Browser, or React Native), and creates projects with proper imports and credentials configuration already in place.&lt;/p&gt;
&lt;h3&gt;Final Thoughts 🤔&lt;/h3&gt;
&lt;p&gt;It&#39;s getting a lot easier to build systems that don&#39;t forget.&lt;/p&gt;
&lt;p&gt;Serverless isn&#39;t stateless anymore. Agents are getting access to real tools, real workflows, and persistent memory. And the infrastructure is finally starting to reflect that shift, with better primitives for state, orchestration, and long-running execution.&lt;/p&gt;
&lt;p&gt;But the tradeoffs are changing. We&#39;re layering memory into systems that were designed to be ephemeral. Giving agents persistence across sessions. Letting them interact with tools and data in ways that blur the line between request and workflow. The hard part isn&#39;t adding memory. It&#39;s deciding what gets promoted from a single session into something durable, what stays scoped to one agent versus shared across many, and what should be forgotten on purpose.&lt;/p&gt;
&lt;p&gt;That&#39;s where things get complicated. State introduces responsibility. Memory introduces risk. Every piece of context an agent carries forward is something you now have to govern. Who can read it, when it expires, how it&#39;s surfaced back into a prompt, and what happens when it&#39;s wrong. The more capable these systems become, the more those decisions start to look like product decisions, not implementation details.&lt;/p&gt;
&lt;p&gt;At the same time, the direction is forming. Serverless platforms are adding stateful primitives. Agent frameworks are focusing on orchestration instead of just prompts. Memory is becoming a first-class concept instead of a bolted-on feature. Even model providers are starting to expose more control over how context is stored, retrieved, and applied.&lt;/p&gt;
&lt;p&gt;It&#39;s not just about generating better responses anymore. It&#39;s about building systems that can carry context forward. Systems that can act, adapt, and remember without breaking the guarantees we still rely on. The shift is real, and it&#39;s accelerating.&lt;/p&gt;
&lt;p&gt;Because once systems start remembering, everything else has to change with them.&lt;/p&gt;
&lt;p&gt;See you next week,&lt;br /&gt;
Jeremy&lt;/p&gt;
&lt;hr /&gt;
&lt;p&gt;I hope you enjoyed this newsletter. We&#39;re always looking for ideas and feedback to make it better and more inclusive, so please feel free to reach out to me via &lt;a href=&quot;https://bsky.app/profile/jeremydaly.com&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Bluesky&lt;/a&gt;, &lt;a href=&quot;http://www.linkedin.com/in/jeremydaly&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;LinkedIn&lt;/a&gt;, &lt;a href=&quot;http://twitter.com/jeremy_daly&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;X&lt;/a&gt;, or &lt;a href=&quot;mailto:contact@jeremydaly.com&quot;&gt;email&lt;/a&gt;.&lt;/p&gt;
</content>
  </entry>
  <entry>
    <title>Issue #362: Mo’ Models, Mo’ Problems ⚠️</title>
    <link href="https://offbynone.io/issues/362/"/>
    <updated>2026-04-21T12:00:00Z</updated>
    <summary>In this issue, Claude gets a major upgrade, AWS makes AI costs more visible, and Cloudflare goes all-in on agents.</summary>
    <id>https://offbynone.io/issues/362/</id>
    <content type="html">&lt;h2&gt;Mo’ Models, Mo’ Problems ⚠️&lt;/h2&gt;
&lt;p&gt;In our &lt;a href=&quot;https://offbynone.io/issues/361&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;previous issue&lt;/a&gt;, AI started breaking things faster than we can defend them, AWS launched an agent registry, and S3 kinda became a filesystem. This week, Claude gets a major upgrade, AWS makes AI costs more visible, and Cloudflare goes all-in on agents. Plus, we&#39;ve got some amazing cloud, serverless, and AI content from the community.&lt;/p&gt;
&lt;h3&gt;News &amp;amp; Announcements&lt;/h3&gt;
&lt;p&gt;Anthropic announced &lt;a href=&quot;https://www.anthropic.com/news/claude-opus-4-7?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23362&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Claude Opus 4.7&lt;/a&gt; this past week as their latest push towards world domination. Early signals point to serious gains in software engineering, especially for long-running tasks, plus stronger vision support. AWS wasted no time &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/04/claude-opus-4.7-amazon-bedrock?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23362&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;rolling it out in Bedrock&lt;/a&gt;.&lt;/p&gt;
&lt;p&gt;AWS also introduced &lt;a href=&quot;https://aws.amazon.com/blogs/machine-learning/introducing-granular-cost-attribution-for-amazon-bedrock?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23362&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;granular cost attribution for Amazon Bedrock&lt;/a&gt;, which is a big step toward actually understanding AI spend. Cost control and observability for LLMs is still pretty messy, and being able to map usage down to IAM users and roles starts to make that problem a lot more tractable.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/04/aurora-serverless-smarter-scaling?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23362&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon Aurora Serverless&lt;/a&gt; is getting up to 30% better performance with smarter scaling, while still keeping the scale-to-zero promise. There’s a deeper dive from the team &lt;a href=&quot;https://aws.amazon.com/blogs/database/aurora-serverless-faster-performance-enhanced-scaling-and-still-scales-down-to-zero?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23362&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;here&lt;/a&gt; if you want the details. I like this direction.&lt;/p&gt;
&lt;p&gt;AWS also announced general availability of &lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/04/aws-announces-ga-AWS-interconnect-multicloud?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23362&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AWS Interconnect&lt;/a&gt;, kicking things off with Google Cloud. Dedicated bandwidth between clouds is becoming a thing, with Azure and Oracle Cloud Infrastructure expected to follow later this year. Let the homogeneity begin.&lt;/p&gt;
&lt;p&gt;Anthropic introduced &lt;a href=&quot;https://claude.com/blog/introducing-routines-in-claude-code?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23362&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;routines in Claude Code&lt;/a&gt;, which basically turns repeatable development workflows into something you can automate. Feels like another positive step toward making agents more useful in day-to-day dev work. They also highlighted what people are building in their ecosystem with &lt;a href=&quot;https://claude.com/blog/meet-the-winners-of-our-built-with-opus-4-6-claude-code-hackathon?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23362&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;their latest hackathon winners&lt;/a&gt;. No fluff, they&#39;re all practical AI solutions that address real pain points. 🤷&lt;/p&gt;
&lt;p&gt;It was Agents Week over at Cloudflare last week, and they shipped &lt;em&gt;a lot&lt;/em&gt;. The full rundown of launches is &lt;a href=&quot;https://blog.cloudflare.com/agents-week-in-review?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23362&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;here&lt;/a&gt;, but there were a few standouts: &lt;a href=&quot;https://blog.cloudflare.com/ai-search-agent-primitive?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23362&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;AI Search&lt;/a&gt; as a core primitive for agents, &lt;a href=&quot;https://blog.cloudflare.com/flagship?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23362&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Flagship&lt;/a&gt; to bring feature flags into the agent era, &lt;a href=&quot;https://blog.cloudflare.com/introducing-agent-memory?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23362&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Agent Memory&lt;/a&gt;, and a new &lt;a href=&quot;https://blog.cloudflare.com/email-for-agents?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23362&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;email service for agents&lt;/a&gt; in public beta.&lt;/p&gt;
&lt;p&gt;Not on my 2026 Bingo card, but Apple announced that &lt;a href=&quot;https://www.apple.com/newsroom/2026/04/tim-cook-to-become-apple-executive-chairman-john-ternus-to-become-apple-ceo?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23362&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Tim Cook is stepping into the Executive Chairman role at Apple, with John Ternus taking over as CEO&lt;/a&gt;. Big shift for one of the most stable leadership runs in tech. I&#39;m sure it has nothing to do with Apple Intelligence. 😬&lt;/p&gt;
&lt;p&gt;And in case you missed it, the recent &lt;a href=&quot;https://vercel.com/kb/bulletin/vercel-april-2026-security-incident?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23362&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Vercel hack&lt;/a&gt; highlights a growing pattern in cybersecurity. Third-party AI tooling accessing internal systems is introducing a whole new threat model. One that most teams aren’t even aware of, never mind prepared for.&lt;/p&gt;
&lt;p&gt;If your incident response still involves five tabs, three tools, and someone asking “who’s on point?”, it might be time to rethink things. &lt;a href=&quot;https://fandf.co/3OcoQib&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;incident.io&lt;/a&gt; is an all-in-one platform that runs Slack and Teams native, so you can declare, manage, and resolve incidents without leaving the conversation. It handles the busywork too, auto-assigning roles, kicking off workflows, and even surfacing insights from past incidents so you don’t keep fixing the same problem twice. Definitely worth a deeper look if you want faster response times without adding more process: &lt;a href=&quot;https://fandf.co/3OcoQib&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;incident.io&lt;/a&gt;. &lt;code&gt;Sponsored&lt;/code&gt;&lt;/p&gt;
&lt;h3&gt;Tutorials&lt;/h3&gt;
&lt;ul&gt;
&lt;li&gt;&lt;a href=&quot;https://edjgeek.com/blog/lambda-managed-instances?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23362&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Lambda Managed Instances: A Working Demo and the Math Behind It&lt;/a&gt; by Eric Johnson&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/machine-learning/transform-retail-with-aws-generative-ai-services?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23362&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Transform retail with AWS generative AI services&lt;/a&gt; by Bhavya Chugh&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://dev.to/jayaganeshk/the-hidden-cost-of-aws-lambda-snapstart-for-python-and-how-i-fixed-it-with-durable-functions-2ba4?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23362&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;The Hidden Cost of AWS Lambda SnapStart for Python, and How I Fixed It with Durable Functions&lt;/a&gt; by Jaya Ganesh&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://claude.com/blog/best-practices-for-using-claude-opus-4-7-with-claude-code?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23362&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Best practices for using Claude Opus 4.7 with Claude Code&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/machine-learning/power-video-semantic-search-with-amazon-nova-multimodal-embeddings?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23362&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Power video semantic search with Amazon Nova Multimodal Embeddings&lt;/a&gt; by Amit Kalawat&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://dev.to/aws-heroes/serverless-applications-on-aws-with-lambda-using-java-25-api-gateway-and-dynamodb-part-6-using-1ji?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23362&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Serverless applications on AWS with Lambda using Java 25, API Gateway and DynamoDB - Part 6 Using GraalVM Native Image&lt;/a&gt; by Vadym Kazulkin&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/database/accelerate-database-migration-to-amazon-aurora-dsql-with-kiro-and-amazon-bedrock-agentcore?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23362&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Accelerate database migration to Amazon Aurora DSQL with Kiro and Amazon Bedrock AgentCore&lt;/a&gt; by Noorul Mahajabeen Mustafa&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://www.infoq.com/articles/lambda-extension-deferred-flush?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23362&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Using AWS Lambda Extensions to Run Post-Response Telemetry Flush&lt;/a&gt; by Melvin Philips&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://dev.to/aws-heroes/serverless-applications-on-aws-with-lambda-using-java-25-api-gateway-and-aurora-dsql-part-5-3dlj?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23362&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Serverless applications on AWS with Lambda using Java 25, API Gateway and Aurora DSQL - Part 5 SnapStart + full priming&lt;/a&gt; by Vadym Kazulkin&lt;/li&gt;
&lt;/ul&gt;
&lt;h3&gt;Reads&lt;/h3&gt;
&lt;p&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/machine-learning/navigating-the-generative-ai-journey-the-path-to-value-framework-from-aws?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23362&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Navigating the generative AI journey: The Path-to-Value framework from AWS&lt;/a&gt;&lt;br /&gt;
AWS tries to put some structure around the chaos with a “Path-to-Value” framework. It’s less of a step-by-step guide and more of a reminder that AI adoption is messy, multidimensional, and mostly about tradeoffs between value, risk, and organizational reality.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://blog.cloudflare.com/past-bots-and-humans?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23362&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Moving past bots vs. humans&lt;/a&gt;&lt;br /&gt;
The bot vs human model is breaking down fast. Cloudflare is leaning into intent over identity, which feels like the right direction as agents start acting more like users and users start looking more like bots.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://blog.cloudflare.com/internal-ai-engineering-stack?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23362&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;The AI engineering stack we built internally — on the platform we ship&lt;/a&gt;&lt;br /&gt;
Always interesting when a company dogfoods its own stack at scale. Cloudflare’s setup is a good look at what a modern AI platform actually needs when you’re pushing billions of tokens and not just running demos.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://www.taskade.com/blog/multi-agent-production?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23362&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Multi-Agent AI in Production | Taskade Engineering (2026)&lt;/a&gt;&lt;br /&gt;
Three years into multi-agent systems and the same problems keep showing up. Memory, coordination, and agents getting stuck in loops. Good practical patterns here, especially if you’ve already hit these walls.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://dev.to/awshuss/why-aws-certified-genai-developer-stands-apart-from-other-aws-certs-14n?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23362&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Why AWS Certified GenAI Developer stands apart from other AWS certs&lt;/a&gt;&lt;br /&gt;
Anwaar Hussain points out that this cert is less about knowing AI and more about wiring it into real systems. Which is probably the right shift, because building with AI is quickly becoming more of an architecture problem than a modeling one.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://dev.to/amitkayal/lessons-i-learned-building-a-memory-aware-agent-with-amazon-bedrock-agentcore-runtime-4lc9?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23362&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Lessons I learned building a memory-aware agent with Amazon Bedrock AgentCore Runtime&lt;/a&gt;&lt;br /&gt;
Memory is still the hardest part of agent design. Amit Kayal gives us a solid walkthrough of scoping, lifecycle, and not blowing up your prompts while trying to make agents feel stateful.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://newsletter.pragmaticengineer.com/p/learnings-from-conducting-1000-interviews?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23362&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Learnings from conducting ~1,000 interviews at Amazon&lt;/a&gt;&lt;br /&gt;
Steve Huynh shares a good reminder that hiring is its own system with its own signals. If you don’t understand what a company actually optimizes for, you’re probably optimizing for the wrong thing.&lt;/p&gt;
&lt;h3&gt;Podcasts, Videos, and more&lt;/h3&gt;
&lt;p&gt;&lt;a href=&quot;https://www.youtube.com/watch?v=SvKXhFVVbGY?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23362&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Serverless Apache Airflow | Serverless Office Hours&lt;/a&gt;&lt;br /&gt;
Airflow, but make it serverless. John Jackson and Kamen Sharlandjiev breakdown when MWAA actually makes sense versus just reaching for Step Functions, especially once you factor in cost, scaling, and how much orchestration complexity you really need.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://www.youtube.com/watch?t=3s&amp;v=xqRUnoaQiUM?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23362&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Building, Managing &amp;amp; Governing APIs on AWS&lt;/a&gt;&lt;br /&gt;
APIs aren’t just for humans anymore. Giedrius Praspaliauskas covers the full lifecycle on AWS, but the interesting part is how API strategies are evolving to support agents, not just apps. Same primitives, very different consumers.&lt;/p&gt;
&lt;h3&gt;New from AWS&lt;/h3&gt;
&lt;ul&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/04/amazon-msk-replicator-external-kafka-cluster-support?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23362&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;MSK Replicator now supports replication from external Apache Kafka clusters to MSK Express Brokers&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/04/amazon-msk-replicator-enhanced-consumer-offset-synchronization?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23362&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon MSK Replicator now supports enhanced consumer offset synchronization for bidirectional replication&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/04/amazon-msk-replicator-logs?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23362&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon MSK Replicator now supports log forwarding for replication visibility&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/04/amazon-documentdb-mongodb-in-place-version-upgrade-5-0-to-8-0?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23362&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon DocumentDB (with MongoDB compatibility) now supports in-place upgrade from version 5.0 to 8.0&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/04/amazon-ecr-pull-through-cache-referrers?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23362&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon ECR Pull Through Cache Now Supports Referrer Discovery and Sync&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&quot;https://aws.amazon.com/about-aws/whats-new/2026/04/amazon-cloudwatch-cross-region-enablement-rules?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23362&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Amazon CloudWatch now supports cross-region telemetry auditing and enablement rules&lt;/a&gt;&lt;/li&gt;
&lt;/ul&gt;
&lt;h3&gt;Developer Tools&lt;/h3&gt;
&lt;p&gt;&lt;a href=&quot;https://github.com/brognilucas/sls-testing?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23362&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;brognilucas/sls-testing&lt;/a&gt; by Lucas Brogni&lt;br /&gt;
Typed, composable testing utilities for AWS Lambda from Lucas Brogni that provides event builders and Jest matchers for Lambda functions.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://dev.to/pujaaan/i-got-tired-of-writing-the-same-cdk-wiring-so-i-built-simple-cdk-obg?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23362&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;pujaaan/simple-cdk&lt;/a&gt; by pujaaan&lt;br /&gt;
A thin runtime over CDK that scans your folders, runs adapters in a deterministic three-phase pipeline (discover → register → wire), and emits real CDK constructs.&lt;/p&gt;
&lt;p&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/machine-learning/toolsimulator-scalable-tool-testing-for-ai-agents?utm_source=newsletter&amp;utm_medium=email&amp;utm_content=offbynone&amp;utm_campaign=Off-by-none%3A%20Issue%20%23362&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;ToolSimulator: scalable tool testing for AI agents&lt;/a&gt; by Darren Wang&lt;br /&gt;
An LLM-powered tool simulation framework within Strands Evals to thoroughly and safely test AI agents that rely on external tools, at scale.&lt;/p&gt;
&lt;h3&gt;Final Thoughts 🤔&lt;/h3&gt;
&lt;p&gt;It’s getting a lot easier to build powerful systems.&lt;/p&gt;
&lt;p&gt;Models are getting better at real work. Agents are starting to handle meaningful workflows. And the infrastructure around all of this is finally catching up, from cost visibility to orchestration to deployment patterns.&lt;/p&gt;
&lt;p&gt;But the gaps are still there.&lt;/p&gt;
&lt;p&gt;We’re wiring these capabilities into systems that were never designed for autonomous behavior. Giving tools access to internal systems. Letting agents make decisions across boundaries that used to be tightly controlled. And in some cases, we’re doing it faster than we understand the implications.&lt;/p&gt;
&lt;p&gt;That’s where things start to break.&lt;/p&gt;
&lt;p&gt;The Vercel incident isn’t an outlier. It’s a preview. A glimpse into what happens when powerful models meet loosely defined boundaries and third-party integrations. The tooling is evolving quickly, but the assumptions behind our systems haven’t fully caught up yet.&lt;/p&gt;
&lt;p&gt;At the same time, you can see the industry starting to respond.&lt;/p&gt;
&lt;p&gt;Better cost attribution. More structured agent workflows. Dedicated primitives for memory, search, and control. Even multicloud connectivity is starting to blur the lines between platforms. It’s not just about building faster anymore, it’s about building systems that can actually support what we’re asking them to do.&lt;/p&gt;
&lt;p&gt;Still early. Still messy. But the pattern is emerging. More power, more abstraction, and more responsibility to get the boundaries right.&lt;/p&gt;
&lt;p&gt;Because “Mo’ Models, Mo’ Problems” isn’t really a joke. It’s just the beginning.&lt;/p&gt;
&lt;p&gt;See you next week,&lt;br /&gt;
Jeremy&lt;/p&gt;
&lt;hr /&gt;
&lt;p&gt;I hope you enjoyed this newsletter. We&#39;re always looking for ideas and feedback to make it better and more inclusive, so please feel free to reach out to me via &lt;a href=&quot;https://bsky.app/profile/jeremydaly.com&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;Bluesky&lt;/a&gt;, &lt;a href=&quot;http://www.linkedin.com/in/jeremydaly&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;LinkedIn&lt;/a&gt;, &lt;a href=&quot;http://twitter.com/jeremy_daly&quot; target=&quot;_blank&quot; rel=&quot;noopener&quot;&gt;X&lt;/a&gt;, or &lt;a href=&quot;mailto:contact@jeremydaly.com&quot;&gt;email&lt;/a&gt;.&lt;/p&gt;
</content>
  </entry>
</feed>