MCP is going stateless: What the new spec means for AI agents
MCP's next spec revision goes stateless — New Relic breaks down what that shift means for agents built against today's session model.
14 articles · 5 categories
The finishable daily brief
Friday, Jul 31, 2026
14 articles · 5 categories
read top to bottom · then stop
In 30 seconds
Agent tooling kept moving from spec to production today: MCP's next revision goes stateless, Dropbox is already using it to pipe security context into AI code reviews, and LangChain built a benchmark to score code-review agents against real reviewer feedback instead of synthetic tests.
China's open-weight momentum is now as much an infrastructure story as a model one — DeepSeek is building a 1-gigawatt data center in Inner Mongolia, and Kimi K3 hit 41% of global open-source downloads within two days of launch, straining Moonshot's own serving capacity.
MCP keeps hardening for production use: its next spec revision drops server-side state, and Dropbox is already using it to surface security context inside AI code reviews. Alongside that, Amazon Quick shipped a natural-language agent for cataloging data assets, and LangChain built ReviewBench to score code-review agents against real reviewer feedback.
MCP's next spec revision goes stateless — New Relic breaks down what that shift means for agents built against today's session model.
Dropbox wired MCP into its internal knowledge platform Dash so AI code-review tools can pull threat models and security requirements straight into the review.
Amazon Quick's new Agentic Catalog Experience, now in preview, lets data curators describe what they need in natural language and auto-generates Datasets and Topics with inherited semantics.
LangChain built ReviewBench, a benchmark that scores code-review agents against real PR feedback from trusted human reviewers instead of synthetic tests.
Two solo-built Show HN launches point at the same instinct: replace legacy productivity formats with ones an AI can write directly, from a local-first knowledge OS to a Markdown-based slide format.
Show HN: Brainstorm is an open-source, local-first “AI-native OS” for personal knowledge management, shipped after three months of solo development.
Show HN: Slaide is an open-source Markdown slide format that AI can write directly and PowerPoint can still open, built to kill manual deck-building.
Model economics moved on two fronts today: DeepSeek's retrained smaller model is now beating its own flagship on agent benchmarks, and OpenAI's GPT-5.6 price cuts point to intelligence getting cheaper faster than headline model releases suggest.
DeepSeek's retrained V4-Flash — its smaller, faster tier — now beats flagship V4-Pro across nine separate agent benchmarks, per Tech Times.
GPT-5.6 pricing dropped 20–80% versus GPT-5.4, and Latent Space pegs the cost of equivalent intelligence as down 13x in four months.
China's open-weight push is now as much an infrastructure story as a model story: DeepSeek is building gigawatt-scale compute, and Kimi K3 is straining serving capacity two days after launch while analysts argue its real moat is the infra behind it, not the open weights.
DeepSeek is building a 1-gigawatt AI data center in Inner Mongolia, Briefs Finance reports — compute capacity on par with major hyperscaler buildouts.
Pandaily argues Moonshot's real moat with Kimi K3 isn't the open model weights but the infrastructure engineering behind serving it, which is much harder to copy.
Kimi K3 hit 41% of global open-source model downloads within two days of launch and grew paid overseas users 4x, per Pandaily, overwhelming Moonshot's serving capacity.
Security framing shifted up a level today: Google Cloud's CISO office is pitching AI threat defense as a board-level agenda item, while OpenAI published both its EU AI Act compliance practices and a takedown of a Cambodia-based scam ring that abused ChatGPT.
Google Cloud's CISO office argues AI threat defense has become a board-level agenda item, not just a security team concern, in the second installment of its monthly Cloud CISO Perspectives series.
OpenAI published its safety, security, transparency, and provenance practices for Europe, framed explicitly around EU AI Act compliance as the law rolls out.
OpenAI disrupted a Cambodia-based scam ring that was using ChatGPT to run investment, romance, gambling, and impersonation schemes at scale.
You are caught up for this edition