{"week":"2026-W34","start":"2026-08-15","end":"2026-08-21","title":"What happened in AI — Aug 15–21, 2026","generated_at":"2026-08-21T21:15:25Z","intro":["Money moved toward inference infrastructure this week, not model training: Stripe bought OpenRouter for $7B and NVIDIA reverse-execuhired Poolside for $12B, while memory prices climbed 500% in a year and Gartner projects agentic-workflow inference costs will more than quintuple by 2028.","That cost pressure is exactly why open weights matter: Qwen 3.8 27B tied GPT-5.6 Luna's score on the Artificial Analysis Intelligence Index, and GLM-5.3 pushed Zhipu's post-training scaling law forward, both giving builders a cheaper lever to pull than routing every call to a frontier API.","Anthropic and OpenAI answered with enterprise proof points — Asana cut five years of engineering work to two weeks with Codex, Anthropic took computer use, Skills, and Files APIs to general availability — while the EU AI Act's watermarking mandate took effect and AWS, Cloudflare, and independent researchers shipped new controls for agent tool-call risk."],"highlights":["Stripe acquires OpenRouter for $7B; NVIDIA's $12B reverse-execuhire sends Poolside's founders and 7GW of neocloud capacity to Nvidia.","Memory prices are up 500% in 12 months, and Gartner projects agentic-workflow inference costs will more than quintuple by 2028.","Qwen 3.8 27B ties GPT-5.6 Luna at 52 on the Artificial Analysis Intelligence Index; GLM-5.3 advances Zhipu's post-training scaling law.","Anthropic took computer use, the Skills API, and the Files API to general availability, while Anthropic and OpenAI both published a wave of enterprise case studies (monday.com, Slack, ABC Legal, Stampli, Asana).","The EU AI Act's Article 50 watermarking mandate takes effect, and Cloudflare (WriteGuard) and AWS (Dogwood) both shipped new security controls for MCP servers and agent tool-call sequences."],"article_count":37,"categories":[{"name":"Money Chasing Compute","slug":"money-chasing-compute","summary":"Big-money deals and rising costs converged on one asset this week: inference capacity. Stripe bought OpenRouter for $7B, NVIDIA reverse-execuhired Poolside for $12B, and memory prices climbed 500% in a year — signs that compute, not model weights, is where the money is moving.","articles":[{"title":"[AINews] Stripe buys OpenRouter for $7B","summary":"Stripe acquired the AI model-routing platform for $7B, betting on distribution and infrastructure rather than building its own models.","source":"latent_space","url":"https://www.latent.space/p/ainews-stripe-buys-openrouter-for","published":"2026-08-17T23:13:41Z"},{"title":"[AINews] Poolside gets $12B reverse-execuhire to NVIDIA; founders stay for $1B, employees go for $6B, Infraco scaling to 7GW neocloud","summary":"NVIDIA's reverse-execuhire sends Poolside's founders over for $1B and its employees for $6B, while its Infraco scales toward 7GW of neocloud capacity.","source":"latent_space","url":"https://www.latent.space/p/ainews-poolside-gets-12b-reverse","published":"2026-08-21T05:45:21Z"},{"title":"[AINews] Memory prices up 500% in 12 months","summary":"DRAM and memory prices have risen 500% in a year, a crunch Latent Space likens to Moore's Law running in reverse back to 2007 levels.","source":"latent_space","url":"https://www.latent.space/p/ainews-memory-prices-up-500-in-12","published":"2026-08-19T08:44:52Z"},{"title":"Inference Costs per Agentic Workflow to Increase More Than Fivefold Through 2028","summary":"Gartner projects agentic-workflow inference costs will more than quintuple by 2028 as agent adoption scales.","source":"hackernews_ai","url":"https://www.gartner.com/en/newsroom/press-releases/2026-08-17-gartner-predicts-ai-inference-costs-per-agentic-workflow-will-increase-more-than-fivefold-through-2028","published":"2026-08-18T17:45:49Z"},{"title":"Frontier Model Cost and Open-Weights Popularity is Driving Demand for Model Routing","summary":"Glean's CEO explains why rising frontier-model costs and stronger open weights are pushing enterprises toward model routing to control spend.","source":"latent_space","url":"https://www.latent.space/p/glean-model-routing","published":"2026-08-18T21:41:10Z"},{"title":"Teaching Everyone to Fish for Tokens","summary":"Nvidia's pitch to enterprises has shifted to build-and-own-your-model rather than buy inference from Anthropic or OpenAI.","source":"interconnects","url":"https://www.interconnects.ai/p/teaching-everyone-to-fish-for-tokens","published":"2026-08-17T15:07:49Z"},{"title":"Agentic AI has made CPUs the new performance bottleneck","summary":"As agent workloads push more general-purpose compute, CPUs — not just GPUs — are becoming the constraint on agentic AI throughput.","source":"hackernews_ai","url":"https://spectrum.ieee.org/ai-cpu-comeback","published":"2026-08-20T01:14:05Z"},{"title":"Securing the Infrastructure of Intelligence","summary":"NVIDIA frames AI factories as the defining infrastructure of the AI era, where compute increasingly functions as a direct revenue driver rather than a cost center.","source":"nvidia_blog","url":"https://blogs.nvidia.com/blog/securing-the-infrastructure-of-intelligence/","published":"2026-08-17T12:34:51Z"}]},{"name":"Open Models Keep Closing the Gap","slug":"open-models-keep-closing-the-gap","summary":"Qwen 3.8 27B and GLM-5.3 posted intelligence-index scores within a point of frontier proprietary models this week, reinforcing that permissively-licensed open weights are now a credible substitute for paid frontier APIs on many tasks.","articles":[{"title":"Qwen 3.8 27B scores 52 on the Artificial Analysis Intelligence Index","summary":"Alibaba's Apache-2-licensed 27B vision model ties GPT-5.6 Luna's score and lands one point behind GLM-5.2 and DeepSeek V4 Pro, both far larger models.","source":"simon_willison","url":"https://simonwillison.net/2026/Aug/17/qwen-38-27b-scores-52/","published":"2026-08-17T23:58:14Z"},{"title":"Qwen 3.8 27B is excellent, but it defaults to wildly overthinking things","summary":"Simon Willison's hands-on take: the 27B model is an ideal size to run locally, but reasons far longer than needed on simple prompts by default.","source":"simon_willison","url":"https://simonwillison.net/2026/Aug/16/qwen-38-27b/","published":"2026-08-16T22:00:39Z"},{"title":"[AINews] Death of Params: Z.ai CEO Jie Tang on GLM 5.3 and the new Post-training Scaling Law","summary":"Zhipu's CEO argues post-training technique, not raw parameter count, is now the scaling law driving GLM-5.3's gains.","source":"latent_space","url":"https://www.latent.space/p/ainews-death-of-params-zai-ceo-jie","published":"2026-08-20T05:17:12Z"},{"title":"NVIDIA Nemotron 3.5 Lightning now available in Amazon SageMaker JumpStart","summary":"The 30B mixture-of-experts model (3B active), built for high-volume agentic workloads, is now deployable directly from SageMaker JumpStart.","source":"aws_ml_blog","url":"https://aws.amazon.com/blogs/machine-learning/nvidia-nemotron-3-5-lightning-now-available-in-amazon-sagemaker-jumpstart/","published":"2026-08-17T18:06:33Z"}]},{"name":"Anthropic and OpenAI Push Agents Into the Enterprise","slug":"anthropic-openai-enterprise-agents","summary":"Both labs spent the week proving agent ROI with concrete numbers — Asana cut five years of engineering work to two weeks with Codex, Stampli cut launch hours 68% with ChatGPT Work — while Anthropic shipped computer use, Skills, and Files APIs to general availability.","articles":[{"title":"Build production agents with computer use, the Skills API, and the Files API | Claude by Anthropic","summary":"Anthropic took computer use, the Skills API, and the Files API to general availability on the Claude Platform, adding a browser-use tool for agents that work inside web applications.","source":"claude_blog","url":"https://claude.com/blog/computer-use-skills-api-files-api","published":"2026-08-20T00:00:00Z"},{"title":"The Claude Code Guide For Startups | Claude by Anthropic","summary":"Anthropic distills five rules startups use with Claude Code to ship at 10x their headcount, from everyone-ships culture to AI-native SDLCs.","source":"claude_blog","url":"https://claude.com/blog/claude-code-guide-for-startups","published":"2026-08-20T00:00:00Z"},{"title":"How monday.com transformed its platform into an agent-first product where humans and agents collaborate | Claude by Anthropic","summary":"monday.com rebuilt its platform around Claude so humans and agents collaborate directly inside the product rather than through a bolt-on chatbot.","source":"claude_blog","url":"https://claude.com/blog/how-monday-com-transformed-its-platform-into-an-agent-first-product-where-humans-and-agents-collaborate","published":"2026-08-20T00:00:00Z"},{"title":"Turning conversation into knowledge: how Slack builds human-agent teams | Claude by Anthropic","summary":"Slack's Chief Product Officer describes how the company turns everyday conversation into structured knowledge agents can act on.","source":"claude_blog","url":"https://claude.com/blog/turning-conversation-into-knowledge-how-slack-builds-human-agent-teams","published":"2026-08-19T00:00:00Z"},{"title":"How ABC Legal turned every employee into a builder with Claude Managed Agents | Claude by Anthropic","summary":"ABC Legal moved from scattered AI experiments to a governed fleet of specialized Claude Managed Agents across the organization.","source":"claude_blog","url":"https://claude.com/blog/how-abc-legal-turned-every-employee-into-a-builder-with-claude-managed-agents","published":"2026-08-17T00:00:00Z"},{"title":"Stampli cuts launch hours by 68% using ChatGPT Work","summary":"With a fixed deadline and design resources committed elsewhere, Stampli used Codex and ChatGPT Work to compress weeks of launch work into days, cutting hours by 68%.","source":"openai_blog","url":"https://openai.com/index/stampli","published":"2026-08-20T00:00:00Z"},{"title":"Asana cleared 5 years of engineering work in 2 weeks with Codex","summary":"Asana used Codex to replace an outdated testing system in two weeks — work it estimated would otherwise take five years — for about $12K.","source":"openai_blog","url":"https://openai.com/index/asana","published":"2026-08-18T07:00:00Z"},{"title":"Replit expands access to software creation with GPT-5.6 Luna","summary":"Replit's new Free Mode, powered by GPT-5.6 Luna, lets anyone turn ideas into working software without worrying about token costs.","source":"openai_blog","url":"https://openai.com/index/replit","published":"2026-08-19T07:00:00Z"}]},{"name":"AI Policy, Safety, and Agent Security","slug":"ai-policy-safety-agent-security","summary":"Regulators and platform teams both moved on agent risk this week: the EU AI Act's watermarking mandate took effect, and AWS, Cloudflare, and independent researchers all shipped new tools and benchmarks for containing what autonomous agents can do.","articles":[{"title":"Offering Zero Data Retention for frontier models","summary":"OpenAI reaffirmed Zero Data Retention for eligible API customers and previewed Private Safety Processing, aimed at advanced AI safety monitoring without compromising data privacy.","source":"openai_blog","url":"https://openai.com/index/our-commitment-to-zero-data-retention","published":"2026-08-19T19:00:00Z"},{"title":"Pacing model development in an era of cyber-critical capabilities","summary":"OpenAI says it is strengthening monitoring, alignment, and security processes to guide how fast it ships models with cyber-critical capabilities.","source":"openai_blog","url":"https://openai.com/index/pacing-model-development-cyber-capabilities","published":"2026-08-18T11:00:00Z"},{"title":"The Defender’s Window","summary":"OpenAI lays out how AI is reshaping cybersecurity for both attackers and defenders, and what security teams can do now while defenders still have an edge.","source":"openai_blog","url":"https://openai.com/index/the-defenders-window","published":"2026-08-17T05:30:00Z"},{"title":"Major Frontier Model Providers Adopt Watermarking Tech to Comply with EU Regulation","summary":"The EU AI Act's Article 50, in force since August 2, requires AI systems to mark synthetic outputs in a machine-detectable way; major vendors are rolling out statistical watermarking to comply.","source":"infoq_ai_ml","url":"https://www.infoq.com/news/2026/08/eu-ai-content-watermark/?utm_campaign=infoq_content&utm_source=infoq&utm_medium=feed&utm_term=AI%2C+ML+%26+Data+Engineering","published":"2026-08-18T05:05:00Z"},{"title":"AWS Open-Sources Dogwood, Extending Cedar to Govern Sequences of Agent Tool Calls","summary":"Dogwood adds temporal conditions to AWS's Cedar policy language, so rules can reason about an agent's prior tool calls — approvals, rate limits — rather than judging one request in isolation.","source":"infoq_ai_ml","url":"https://www.infoq.com/news/2026/08/aws-dogwood-agent-policy/?utm_campaign=infoq_content&utm_source=infoq&utm_medium=feed&utm_term=AI%2C+ML+%26+Data+Engineering","published":"2026-08-16T07:26:00Z"},{"title":"Cloudflare WriteGuard Brings Fine-Grained Security Controls for MCP Servers","summary":"Now in private beta, WriteGuard gives fine-grained control over which tools an AI agent can access through an MCP server.","source":"infoq_ai_ml","url":"https://www.infoq.com/news/2026/08/cloudflare-writeguard-mcp-safety/?utm_campaign=infoq_content&utm_source=infoq&utm_medium=feed&utm_term=AI%2C+ML+%26+Data+Engineering","published":"2026-08-18T16:00:00Z"},{"title":"An open agent-security benchmark, including the attacks we fail to catch","summary":"An open-source agent-security benchmark ships alongside a public list of attacks the author's own tooling still fails to catch.","source":"hackernews_ai","url":"https://github.com/AndrewSispoidis/contemporary-agent-attacks","published":"2026-08-16T14:38:05Z"},{"title":"SecIT Bench A frontier benchmark for AI agents in IT and security workflows","summary":"A new frontier benchmark evaluates AI agents specifically on IT and security operations tasks.","source":"hackernews_ai","url":"https://secitbench.cribl.io/","published":"2026-08-19T00:28:02Z"}]},{"name":"Agent Engineering Toolbox","slug":"agent-engineering-toolbox","summary":"Eval and observability tooling matured fast this week — Langfuse rebuilt its stack on a single ClickHouse table, LangSmith added tuned evaluators and preview builds — as builders traded early productivity wins for load-bearing infrastructure.","articles":[{"title":"Langfuse v4: agent evals and traces rebuilt on one immutable ClickHouse table","summary":"Langfuse v4 rebuilds its evals and tracing stack on a single immutable ClickHouse table.","source":"hackernews_ai","url":"https://langfuse.com/changelog/2026-08-17-langfuse-v4","published":"2026-08-19T12:46:51Z"},{"title":"Introducing LangSmith Tuned Evaluators","summary":"LangSmith's new Tuned Evaluators attach quality feedback to production traces, starting with a 'Perceived Error' metric to help teams find and fix agent mistakes.","source":"langchain_blog","url":"https://www.langchain.com/blog/introducing-langsmith-tuned-evaluators-starting-with-perceived-error","published":"2026-08-19T12:59:39Z"},{"title":"Test Agent Changes with LangSmith Preview Builds","summary":"Preview Builds let teams test pull-request branches in temporary, production-like LangSmith deployments before merging agent changes.","source":"langchain_blog","url":"https://www.langchain.com/blog/langsmith-preview-builds-test-agent-changes-before-production","published":"2026-08-20T19:20:50Z"},{"title":"AgentCore Payments middleware for LangChain agents","summary":"New middleware lets LangChain agents pay for APIs with deterministic session budgets, signing x402 payments that LangSmith traces automatically.","source":"langchain_blog","url":"https://www.langchain.com/blog/langchain-agentcore-payments","published":"2026-08-18T19:27:37Z"},{"title":"The /wayfinder Skill: Navigating the “Fog of War” of Planning","summary":"Matt Pocock's /wayfinder skill helps coding agents plan on greenfield projects or when the path forward is genuinely unclear.","source":"latent_space","url":"https://www.latent.space/p/wayfinder-skill","published":"2026-08-20T20:59:09Z"},{"title":"How canvases make agentic workflows visible, steerable, and cost-efficient","summary":"GitHub argues chat interfaces lose agent work in the scroll, and shows how visual canvases keep agentic workflows steerable and cost-efficient instead.","source":"github_blog_ai_ml","url":"https://github.blog/ai-and-ml/github-copilot/how-canvases-make-agentic-workflows-visible-steerable-and-cost-efficient/","published":"2026-08-17T16:00:00Z"},{"title":"Recursive Self-Improvement","summary":"Agents can already edit their own tools, skills, and harness; genuine recursive self-improvement still needs a system that can raise its own verifier without capturing it.","source":"philschmid","url":"https://www.philschmid.de/recursive-self-improvement","published":"2026-08-21T00:00:00Z"},{"title":"IsoExec: Unified Execution to Eliminate Trainer-Inference Mismatch in SkyRL","summary":"IsoExec unifies numerical execution across SkyRL's vLLM and Megatron runtimes, cutting the rollout-versus-training logprob difference below 1e-6 on Qwen3.5-35B-A3B with 25% overhead.","source":"vllm_blog","url":"https://vllm.ai/blog/2026-08-21-isoexec","published":"2026-08-21T00:00:00Z"},{"title":"Cloudflare Cuts Astro Github Issues by 85% with AI Agents","summary":"Cloudflare's AI-agent-driven issue triage cut Astro's open GitHub issues by 85%.","source":"infoq_ai_ml","url":"https://www.infoq.com/news/2026/08/cloudflare-astro-ai-agents/?utm_campaign=infoq_content&utm_source=infoq&utm_medium=feed&utm_term=AI%2C+ML+%26+Data+Engineering","published":"2026-08-21T14:09:00Z"}]}]}