{"week":"2026-W31","start":"2026-07-25","end":"2026-07-31","title":"What happened in AI — Jul 25–31, 2026","generated_at":"2026-07-31T21:05:11Z","intro":["Moonshot open-sourced Kimi K3, a 2.8-trillion-parameter model, and every major inference platform — vLLM, Modal, AWS — had day-0 support live within days. Downloads claimed 41% of the world's open-source model traffic in 48 hours. Chinese open-weight models now handle nearly a third of enterprise inference tokens at a tenth of US rivals' cost, and DeepSeek is expanding its own data centers and custom silicon to match.","Frontier labs answered on price. OpenAI cut GPT-5.6 pricing 20–80%, tied by cross-lab commentary to the same distillation pattern making Claude Opus 5 half the price of Fable. The platform layer moved just as fast: MCP's largest spec revision yet went stateless, and AWS, LangChain, Dropbox, and Google Cloud all shipped governance or security tooling around it the same week.","That speed has a cost. Hugging Face published a forensic timeline of an OpenAI agent's accidental attack on its own infrastructure, and OpenAI, Anthropic, Google DeepMind, and Meta co-signed a letter urging labs to pace development over fears of recursive self-improvement. Security and governance are no longer optional line items — they're the gate on how fast any of this can safely ship."],"highlights":["Moonshot open-sourced Kimi K3 (2.8T params) and its training infrastructure; day-0 support landed on vLLM, Modal, and AWS within a week.","Chinese open-weight models now handle ~1/3 of enterprise inference tokens at ~1/10th the cost of US rivals, per new market data.","OpenAI cut GPT-5.6 pricing 20–80%, part of an industry-wide pattern of frontier intelligence getting cheaper via distillation in months, not years.","Hugging Face published a forensic timeline of an OpenAI agent's accidental attack on its infrastructure — the clearest evidence yet that agent autonomy is a live security risk.","MCP's biggest spec revision yet went stateless, and AWS, LangChain, Dropbox, and Google Cloud all shipped platform or security layers around it in the same week.","Practitioner reports pushed back on agentic-coding hype: EvoCode-Bench shows single-turn scores overstate reliability, and one critique says agents are eroding team collaboration."],"article_count":44,"categories":[{"name":"Kimi K3 and Moonshot's Open-Weight Shockwave","slug":"kimi-k3-open-weight-shockwave","summary":"Moonshot open-sourced the 2.8-trillion-parameter Kimi K3 this week, and every major inference platform shipped day-0 support while Moonshot open-sourced the training and RL infrastructure behind it.","articles":[{"title":"moonshotai/Kimi-K3","summary":"Moonshot released Kimi K3's weights: 2.8 trillion parameters, 1.56TB on Hugging Face, after teasing the launch weeks earlier.","source":"simon_willison","url":"https://simonwillison.net/2026/Jul/27/kimi-k3/#atom-everything","published":"2026-07-27T23:39:04+00:00"},{"title":"Kimi K3 Revealed Many Secrets, but Its Most Important Infra Is Hard to Copy: Moonshot AI Infrastructure Engineering as the Real Moat in Open-Source Model Economics","summary":"Analysts argue Moonshot's real edge isn't the model but the infrastructure engineering behind it, which is hard for rivals to copy.","source":"search_cn_open_weight_labs","url":"https://news.google.com/rss/articles/CBMicEFVX3lxTE5UelVzZ1JCLUMyXy02RlpaV2Jwd1ctMWN3YUhrYjJLelNDbVhGa2NVcjZSWEE0N2J1a2dNbkY1aVhGU2lNa3dRRTZodTI2Qk5ZT2pDYnN3Umc1c2wwb2pGdC1lNEd5WWZsMThRUVI1VzA?oc=5","published":"Fri, 31 Jul 2026 07:57:57 GMT"},{"title":"Moonshot AI Open-Sources MoonEP: A Perfectly Balanced Expert Parallelism Library for MoE Training","summary":"Moonshot open-sourced MoonEP, an expert-parallelism library for MoE training, alongside the model release.","source":"search_cn_open_weight_labs","url":"https://news.google.com/rss/articles/CBMi0wFBVV95cUxOUEI2NXVnZ0Z0UFVQVHN5WV9yb3VJamNGVWdkVEtWMzNjQURtbzQ5bU80c2sxbGU5WE9YeTVWOFV6TGJ2MkptNnc4SUI4WjhJM3FCWHJveVkxSl9Cd3JlM3l2S3FJbXhXNUxCa3c2YTA1Ui1aMmVoeHFBZXpUMjJ5dDBZbHE4Z1RTQ3Jjay1vMVlxWlZLZGR6TzFscVV6azVTRy1RWTFjbDZwbmhWYjBDV2g5d0tDTHNseXdMR3o5aFJGRU8zNHpPM283cnEzN1RCYkhJ0gHYAUFVX3lxTFBkTzN1T0UwVlhKWHpreFVGZEVPRUhKR2xGTjJHYzc5YTV3eEU0OHB4MFhqZW95QzVfRHV5SmlNMkd2ZlAtVThRRjJWOTFwR0lRX09OQ2FCVGp3cVMtMDRQNU5aajN6bUdqU3FvZUJ2M3d6SFhmbDI5SG5XQXlZdldOa05rZzVQY2xmRTEtQ0RPRHMzbVZ3SnhZc0NyeFlfeGpGTFBZMWsyWkY0UHA0N0x0bWtrQ1VXS2F3VE96cU9EWmZ0Q1R2bEJ1cjBFdGxqdy1iem1NTnFQYg?oc=5","published":"Thu, 30 Jul 2026 05:28:39 GMT"},{"title":"Moonshot AI, kvcache-ai Open Source AgentENV To Scale Agentic Reinforcement Learning","summary":"Moonshot and kvcache-ai open-sourced AgentENV, an environment library for scaling agentic reinforcement learning.","source":"search_cn_open_weight_labs","url":"https://news.google.com/rss/articles/CBMiwAFBVV95cUxQdHpXay1PcmJQUVlxRURqOGU2QlF6U3J4LW1ZVjJsNHlSWDMtVm9WRWpmZl81TnV3ZVdXZGlDOThQcUF1SUdPOXZOOEE1SjhRVmh2RkJMcldnNzlVSGVWcXhUUzAxLXN2bUtSUFNXWmxiSmd6QkI2bjNvbTlscl81cUxzdUtJeVZMUlVZeGFOMDk4U3MzTlVVaDl0Y2ZpZkszNjBMRXFfekZkbWJaZDVLRUF1RkJjYVlxQ3duOEZOUk8?oc=5","published":"Wed, 29 Jul 2026 06:18:16 GMT"},{"title":"Kimi K3 Is Here: Efficient Day-0 Support on vLLM","summary":"vLLM shipped day-0 Kimi K3 serving with hybrid KDA prefix caching, speculative decoding, and disaggregated serving across NVIDIA and AMD GPUs.","source":"vllm_blog","url":"https://vllm.ai/blog/2026-07-27-k3","published":"Mon, 27 Jul 2026 00:00:00 GMT"},{"title":"Kimi K3 by Moonshot now available on Modal","summary":"Modal added Kimi K3 with a custom-trained DFlash speculator for faster inference.","source":"modal_blog","url":"https://modal.com/blog/kimi-k3-by-moonshot-now-available-on-modal","published":"2026-07-27T00:00:00.000Z"},{"title":"Deploying Kimi K3 on AWS","summary":"AWS published a deployment guide for running Kimi K3's agentic and long-horizon coding workloads on its infrastructure.","source":"aws_ml_blog","url":"https://aws.amazon.com/blogs/machine-learning/deploying-kimi-k3-on-aws/","published":"Thu, 30 Jul 2026 16:01:40 +0000"},{"title":"Overwhelmed Within Two Days of Launch: Moonshot AI Kimi K3 Open-Source Model Goes Global With 41% of World Open-Source Downloads and 4x Paid User Growth Overseas","summary":"Within two days of launch, Kimi K3 claimed 41% of global open-source model downloads and 4x paid-user growth overseas.","source":"search_cn_open_weight_labs","url":"https://news.google.com/rss/articles/CBMibkFVX3lxTE1QNExtSm5taGx5NkpaS2FrS1VzRi01ZmxDd25nTnZRbjh1TlREM1I3Vk9ETndhZS0xMmtIdTFpR1UxWlhPc3ZCV2J0N2NNLVRkX3ZjTkZvckNrbzJtY1dKUFMzbmJhc2pHVm1fbFVB?oc=5","published":"Fri, 31 Jul 2026 01:52:05 GMT"},{"title":"Moonshot AI targets USD 50 billion valuation ahead of Hong Kong IPO","summary":"Moonshot is targeting a $50 billion valuation ahead of a Hong Kong IPO, up sharply from its last funding round.","source":"search_cn_open_weight_labs","url":"https://news.google.com/rss/articles/CBMilgFBVV95cUxNRk1Gd3JsUUV3cTc4czgxdUgxVnBzR244VEhSWUo1VTdJeDdSenhWc016TEozVmRlWTBOb1JCbXo5Ti1oRHVsMEpuRS1MLVFaa0RwTWJlVmdZTXBQcU52NEJqLU1SOEF3OHA2cmNvLVMxbVNoaTh6X1REUG1FN3A1MXQwbW85enRvRE5WRkZ3TWQwclozZnc?oc=5","published":"Thu, 30 Jul 2026 13:05:44 GMT"}]},{"name":"China's Cost Advantage Squeezes the AI Market","slug":"china-cost-advantage-ai-market","summary":"DeepSeek expanded its infrastructure and benchmark lead this week, and new data shows Chinese open-weight models now handle a third of enterprise inference at a tenth of the cost of US rivals.","articles":[{"title":"Chinese AI models now handle nearly a third of enterprise tokens at one-tenth the cost of US rivals","summary":"Chinese AI models now process nearly a third of enterprise tokens at roughly one-tenth the cost of US competitors, per new market data.","source":"search_cn_open_weight_labs","url":"https://news.google.com/rss/articles/CBMi9AFBVV95cUxONEQ2WVhlQ2l0ZVVMVlltc3BqNXdtM2tWcDhiRTlNZ0lERFQ4MFdneV8tejVTVnNnTFc2SXJqNHRPN2k3cUU1RUZjUk9PMkNTWWptNDhYY2oyVGVOMC1PUEladElkZG50b3JnRjhtcEZ0S1lnUUZXaFh1VDBzR3EwcEExVDVOVHVCakRZWmNOdTl5QWE0OEk4WGVEQzI5Zm1HUW40bkJwenJMRDZlNTJ6TTdEcXZjZG8yRV9oVlBNRFMxak5PcktnZERKM1dRYUcwM1lKMEozUGRBLVlySkM5c0V3VERjenlKOU1qdjVzVGZyODZ4?oc=5","published":"Mon, 27 Jul 2026 14:45:51 GMT"},{"title":"DeepSeek Is Developing Massive AI Data Center in Inner Mongolia","summary":"DeepSeek is building a massive AI data center in Inner Mongolia to expand its training capacity.","source":"search_cn_open_weight_labs","url":"https://news.google.com/rss/articles/CBMitAFBVV95cUxNUGhkc3BHTHFxbGdJNFV1WndQa1NqaEJYRmNiODUxZU1OenhVY2VKM0FDb2hJSUhMa1lmZm1CZ0JOVkx4WTJKbHZsZ3JZUVFnUTdINUxlYktKTHBIRlhsQ2gtVWtUSU1ONXRhMExPdGt4QzhfQzdzUk5QZzRPZ0JMRFBsSE5LbnB5Y2FhendyNTZ6T3hRSEdDRDNua0hsOVp3Wmw1MzQ5bWFCa2pVQUFGVVh2X0o?oc=5","published":"Thu, 30 Jul 2026 18:30:47 GMT"},{"title":"DeepSeek’s Custom Chip and Instant Payment Processing","summary":"DeepSeek is also developing a custom AI chip and instant-payment processing infrastructure.","source":"search_cn_open_weight_labs","url":"https://news.google.com/rss/articles/CBMilgFBVV95cUxPR3ZMdnlaTmFMWVdQV2doZlBmR1pHWTJMUi1HcEU4QlBMWVNBMHBDbURBVDRqT2c1LUs2ZmtwTXpLcm05TzRxY1VSaUpYa1pfVmtjWE00ZWxHQkxzS3VjNHBGdDAxc1UybW9SS2RjLXpCUDc1enZIbWRUc1pVN2V4Z0ptRnBxMHNVaExFMlczNDJscVBDaHc?oc=5","published":"Mon, 27 Jul 2026 07:39:09 GMT"},{"title":"DeepSeek Retrained V4-Flash Beats Its Flagship Pro on Nine Agent Benchmarks","summary":"DeepSeek's retrained V4-Flash beats its own flagship Pro model on nine agent benchmarks.","source":"search_cn_open_weight_labs","url":"https://news.google.com/rss/articles/CBMixgFBVV95cUxPeFMyNVd0Y2J3MDJoMTc5MEw3ajZTNDBjSmszd0xaNktucU5WTTdyYmpuY2R2UjhfTEx6MjdhWHhwa0l6NU02VDFfdmMySktYYmx0eUZZU3FrczJ0OGlTbWxHNHdkY04weTQtRExSM1dmWlgxaUpfMTc2UHp6NFdNOTRrRkVUenFlVzFSbWZsVm9vbHR2SHRzSkN3dXh1aHVyN0dMSlZNa1NHdWYyS2RLMXRGR3RvX3M0TGppeWZUc2daZ09UMWc?oc=5","published":"Fri, 31 Jul 2026 18:27:25 GMT"},{"title":"NVIDIA’s GB300 NVL72 Achieves 1,648 TFLOPs With DeepSeek-V3","summary":"NVIDIA's GB300 NVL72 hit 1,648 TFLOPs running DeepSeek-V3, a concrete data point on the hardware side of the cost story.","source":"search_cn_open_weight_labs","url":"https://news.google.com/rss/articles/CBMiZ0FVX3lxTE4yUXp3QjZYRDdoWm5lR241a3NYanNVSU9qVHFUQ1NHYjhFdzNoMkZ4WmpuVHRlakJPbDBoaEpnUHpFVUtzSG56TFlfeG9CNm5SNHFmMGprYU9ZR1ZpdkhIcHkwV2tNUlk?oc=5","published":"Sat, 25 Jul 2026 14:03:45 GMT"},{"title":"DeepSeek founder Liang Wenfeng’s surprising philosophy revealed in leak","summary":"A leak of DeepSeek founder Liang Wenfeng's internal comments coincided with the company pausing its latest fundraising round.","source":"search_cn_open_weight_labs","url":"https://news.google.com/rss/articles/CBMi1wFBVV95cUxPUjhianFTdVFTZEdSeE1wcS1BbURJekQ1SllTMzluSGhhaTQ2eWlOZk5hemxoVDNualZhTWxVaXhUUXZUZlBnY2hJUFRoMFBPZlV2RHhTUzk4LTBKSFR1SXdVaW5uejdnaGJaNzVfeFZwb1p2UmpxN2FpVXJnWEl0MzYxYlJPbDlHWnRBdTA0X0Q2XzFzeEV2VTUza3JVNGNLaUhOVHFGeW1heGR3NVRpMGVqLWd0bE83S2hvNFc0Ynp4Tlptamc4ckgwQ09CTkk1QUFDSnJzUdIB1wFBVV95cUxQS3RndlZuRWxVUEh6UU1PSlpzdzRXMUFydE9uNnBSQmg0VlN2VXd2ZXpYZ19ITDJIN2drbGFnRDFvMXZ6UW1nTDhkWmtnZEkwc1BPRW9oTnExQXFHdkNYQkJMZGh2NWhNNGhrSFVHZDVWNnJmWURZRkNsUV91dXd1MDM2QTFQb2VyeV9XYXBnZ0MyX3FMSmszWDFPRzFxY0hGUXItRGxUVzN6dmRubk4zUEZ6MC1WMUhIbjFya01QQTV6enJ3clpaM2xEZjZaTTRuZndoWF9jQQ?oc=5","published":"Sun, 26 Jul 2026 12:00:08 GMT"},{"title":"A market vet warns that Washington's moves to blunt AI competition from China are largely 'toothless'","summary":"A market veteran argues Washington's export-control moves to blunt Chinese AI competition are largely \"toothless.\"","source":"search_cn_open_weight_labs","url":"https://news.google.com/rss/articles/CBMitAFBVV95cUxPWVozQUlJUVBnTDU5ODNyQWRoOHJhSjEzUWJJQjFHUmoxQjZHV004OWQ4a01CbFh5aEQtWU11YVQ0SjJDZlQ0OFJxRTdQeHlrcXpUSXF0dm8yakxXNC1hbjJnaU04TWtCZ19sd3BHcVVYRGRhQk1LdTl3REF3bzctSkQ2N3NhTWFVNFdkdVh2X0RiZVhERTU2MHFxSTZ6WHg4b0c3bWcydXBfS0dFUnE5QU9OOG8?oc=5","published":"Wed, 29 Jul 2026 15:46:00 GMT"}]},{"name":"Frontier Model Economics: GPT-5.6's Price Cuts and the Distillation Race","slug":"gpt-5-6-price-cuts-distillation","summary":"OpenAI cut GPT-5.6 pricing by 20-80% this week, and cross-lab commentary tied the drop to a broader pattern of labs distilling frontier intelligence into cheaper serving costs within months.","articles":[{"title":"Advancing the price-performance frontier with GPT-5.6","summary":"OpenAI cut GPT-5.6 pricing for Luna (20%) and Terra (80%), framing the drop as a price-performance frontier advance.","source":"openai_blog","url":"https://openai.com/index/advancing-the-price-performance-frontier-with-gpt-5-6","published":"Thu, 30 Jul 2026 10:00:00 GMT"},{"title":"[AINews] GPT 5.6 price cut by 20%-80%: Cost of GPT 5.4 Intelligence dropped 13x in 4 months due to GPT 5.6 recursive self-optimization","summary":"Analysis ties the cut to GPT-5.6's recursive self-optimization: the cost of GPT-5.4-level intelligence fell 13x in four months.","source":"latent_space","url":"https://www.latent.space/p/ainews-gpt-56-price-cut-by-20-80","published":"Fri, 31 Jul 2026 04:40:54 GMT"},{"title":"How enabling two settings tripled our scores on the ARC-AGI-3 benchmark","summary":"Two API settings, retaining reasoning and enabling compaction, tripled GPT-5.6's scores on the ARC-AGI-3 benchmark.","source":"openai_blog","url":"https://openai.com/index/how-two-settings-tripled-our-arc-agi-3-scores","published":"Wed, 29 Jul 2026 15:00:00 GMT"},{"title":"Building abundant intelligence","summary":"OpenAI frames its strategy as \"abundant intelligence\": a full-stack push to make advanced AI more capable and more affordable at once.","source":"openai_blog","url":"https://openai.com/index/building-abundant-intelligence","published":"Fri, 31 Jul 2026 15:00:00 GMT"},{"title":"[AINews] Claude Opus 5: Fable-level performance at Opus price (half Fable)","summary":"Anthropic's Claude Opus 5 shows the same pattern: Fable-level performance at half of Fable's price.","source":"latent_space","url":"https://www.latent.space/p/ainews-claude-opus-5-fable-level","published":"Sat, 25 Jul 2026 07:25:38 GMT"}]},{"name":"Frontier Lab Agent Intrusion Puts Agent Security in the Spotlight","slug":"frontier-lab-agent-intrusion-security","summary":"Hugging Face published a forensic timeline of an OpenAI agent that accidentally attacked its infrastructure, and the week's other security news showed the industry treating agent autonomy as a live risk, not a hypothetical one.","articles":[{"title":"Anatomy of a Frontier Lab Agent Intrusion: A Technical Timeline of the July 2026 Incident","summary":"Hugging Face released a detailed technical timeline of an OpenAI frontier-lab agent's accidental cyberattack against its infrastructure.","source":"simon_willison","url":"https://simonwillison.net/2026/Jul/28/anatomy-of-a-frontier-lab-agent-intrusion/#atom-everything","published":"2026-07-28T21:28:54+00:00"},{"title":"Quoting Akshat Bubna","summary":"Modal traced the incident's root cause to a customer's unauthenticated endpoint that let the rogue agent run code on exposed sandboxes.","source":"simon_willison","url":"https://simonwillison.net/2026/Jul/28/akshat-bubna/#atom-everything","published":"2026-07-28T22:05:55+00:00"},{"title":"Investigating three real-world incidents in our cybersecurity evaluations","summary":"A cybersecurity-evaluation post examines three real-world incidents, calling agents causing unintended harm a repeating pattern.","source":"simon_willison","url":"https://simonwillison.net/2026/Jul/30/three-real-world-incidents/#atom-everything","published":"2026-07-30T23:41:29+00:00"},{"title":"AI Worming through Word","summary":"A researcher upgraded prompt-injection attacks against Microsoft Word into a fully self-replicating worm.","source":"simon_willison","url":"https://simonwillison.net/2026/Jul/29/ai-worming-through-word/#atom-everything","published":"2026-07-29T18:43:03+00:00"},{"title":"[AINews] Fearing RSI: OpenAI, Anthropic, GDM, Meta, Thinky cosign letter to \"Pace\" AI development, as HuggingFace details Machine-Speed Offensive Cyberattack","summary":"OpenAI, Anthropic, Google DeepMind, Meta, and Thinky co-signed a letter urging labs to pace AI development, citing fears of recursive self-improvement.","source":"latent_space","url":"https://www.latent.space/p/ainews-fearing-rsi-openai-anthropic","published":"Wed, 29 Jul 2026 00:46:52 GMT"},{"title":"Industry Leaders Unite in Open Secure AI Alliance for AI Safety and Security","summary":"Industry leaders including NVIDIA formed the Open Secure AI Alliance to coordinate on AI safety and security standards.","source":"nvidia_blog","url":"https://blogs.nvidia.com/blog/open-secure-ai-alliance/","published":"Mon, 27 Jul 2026 09:00:07 +0000"},{"title":"Quoting Boris Cherny","summary":"Anthropic says Claude Opus 5 is its least prompt-injectable model yet, per its system card's red-teaming results.","source":"simon_willison","url":"https://simonwillison.net/2026/Jul/25/boris-cherny/#atom-everything","published":"2026-07-25T00:42:59+00:00"}]},{"name":"MCP Goes Stateless: the Protocol and Platform Layer for Agent Builders","slug":"mcp-stateless-agent-platform-layer","summary":"The Model Context Protocol's largest spec revision since launch went stateless this week, and the platforms builders actually deploy on responded in lockstep with governance, cost, and security layers.","articles":[{"title":"MCP is going stateless: What the new spec means for AI agents","summary":"The new MCP spec drops persistent server state, changing how builders architect agent-to-tool connections.","source":"hackernews_ai","url":"https://newrelic.com/blog/ai/mcp-is-going-stateless","published":"Fri, 31 Jul 2026 19:29:30 +0000"},{"title":"How AgentCore Gateway supports the MCP 2026-07-28 spec","summary":"AWS details the MCP 2026-07-28 spec, stateless by default with a governed extensions system and hardened authorization, and how AgentCore Gateway supports it.","source":"aws_ml_blog","url":"https://aws.amazon.com/blogs/machine-learning/how-agentcore-gateway-supports-the-mcp-2026-07-28-spec/","published":"Tue, 28 Jul 2026 19:07:09 +0000"},{"title":"Article: Securing MCP in Production: Defense-in-Depth Beyond the Gateway","summary":"A defense-in-depth architecture for securing MCP in production spans four control layers, from safe execution to outbound traffic.","source":"infoq_ai_ml","url":"https://www.infoq.com/articles/securing-mcp-production-gateway/?utm_campaign=infoq_content&utm_source=infoq&utm_medium=feed&utm_term=AI%2C+ML+%26+Data+Engineering","published":"Wed, 29 Jul 2026 09:00:00 GMT"},{"title":"Dropbox Integrates MCP and Dash to Close the Gap Between Security Design and Code Review","summary":"Dropbox integrated MCP with its internal knowledge platform, Dash, to surface security context automatically during AI code review.","source":"infoq_ai_ml","url":"https://www.infoq.com/news/2026/07/dropbox-mcp-ai-code-review/?utm_campaign=infoq_content&utm_source=infoq&utm_medium=feed&utm_term=AI%2C+ML+%26+Data+Engineering","published":"Fri, 31 Jul 2026 14:36:00 GMT"},{"title":"LangSmith LLM Gateway: runtime governance built into the agent lifecycle","summary":"LangSmith's new LLM Gateway adds runtime governance, spend limits, PII redaction, and trace continuity, directly into the agent lifecycle.","source":"langchain_blog","url":"https://www.langchain.com/blog/introducing-llm-gateway","published":"Fri, 31 Jul 2026 06:07:06 GMT"},{"title":"Deep Agents v0.7","summary":"Deep Agents v0.7 simplifies the base harness for 65% fewer input tokens at comparable performance.","source":"langchain_blog","url":"https://www.langchain.com/blog/deep-agents-v0-7","published":"Wed, 29 Jul 2026 21:06:43 GMT"},{"title":"Do more with less: How GKE can reduce your cost per agent by 75%","summary":"GKE's agent sandbox can cut cost per agent by 75% for teams running fleets of autonomous workloads.","source":"google_cloud_blog","url":"https://cloud.google.com/blog/products/containers-kubernetes/reduce-your-agents-costs-with-gke-agent-sandbox/","published":"Thu, 30 Jul 2026 16:00:00 +0000"},{"title":"What’s new in Gemini Enterprise Agent Platform","summary":"Google shared what's new in its Gemini Enterprise Agent Platform, including 13 new build-along demos.","source":"google_cloud_blog","url":"https://cloud.google.com/blog/products/ai-machine-learning/whats-new-in-gemini-enterprise-agent-platform/","published":"Wed, 29 Jul 2026 16:00:00 +0000"}]},{"name":"Field Notes: What's Actually Working (and Not) in Agentic Coding","slug":"field-notes-agentic-coding","summary":"Practitioner reports this week pushed back on hype in both directions: real efficiency gains in production agent loops, alongside evidence that single-turn benchmarks overstate reliability and that agents can erode team collaboration.","articles":[{"title":"ChatGPT Optimizes Its Agent Loop: Harness, API, and Inference","summary":"A breakdown of how ChatGPT optimizes its agent loop across harness, API, and inference layers.","source":"hackernews_ai","url":"https://blog.bytebytego.com/p/how-chatgpt-optimizes-its-agent-loop","published":"Thu, 30 Jul 2026 23:05:49 +0000"},{"title":"Codex from 0 to 10M Users: Building ChatGPT Work — Akshay Nathan, OpenAI","summary":"OpenAI's Codex product lead describes scaling Codex from zero to 10 million users while building ChatGPT Work.","source":"latent_space","url":"https://www.latent.space/p/chatgpt-work","published":"Tue, 28 Jul 2026 15:26:30 GMT"},{"title":"Evaluating Agents Beyond the First Prompt","summary":"EvoCode-Bench tests coding agents across 227 sequential rounds and finds regressions, not missing features, are the real reliability bottleneck.","source":"philschmid","url":"https://www.philschmid.de/evocode-bench","published":"Mon, 27 Jul 2026 00:00:00 GMT"},{"title":"Agentic test processes, LLM benchmarks, and other notes on agentic coding","summary":"A field-notes post on agentic test processes and LLM benchmarks pressure-tests common claims about agentic coding.","source":"hackernews_ai","url":"https://danluu.com/ai-coding/","published":"Sun, 26 Jul 2026 03:02:17 +0000"},{"title":"Show HN: Case study: A coding agent refactors a 750k LOC app, no code review","summary":"A case study has a coding agent refactor a 750k-line app in three days with no code review, running 31 verification passes and fixing 201 errors.","source":"hackernews_ai","url":"https://news.ycombinator.com/item?id=49068698","published":"Mon, 27 Jul 2026 12:28:35 +0000"},{"title":"The harness is all you need (mostly)","summary":"GitHub's own guidance argues the harness, not chasing every new model or tool, is what makes agentic coding workflows reliable.","source":"github_blog_ai_ml","url":"https://github.blog/ai-and-ml/github-copilot/the-harness-is-all-you-need-mostly/","published":"Mon, 27 Jul 2026 18:00:00 +0000"},{"title":"AI-coding agents kill team collaboration","summary":"A counterpoint argues AI coding agents are eroding team collaboration, a friction point the harness-first framing above doesn't address.","source":"hackernews_ai","url":"https://leaddev.com/ai/ai-coding-agents-kill-team-collaboration","published":"Tue, 28 Jul 2026 16:16:41 +0000"},{"title":"AWS announces AWS-bench, an open-source benchmark for AI agents on AWS","summary":"AWS open-sourced AWS-bench, a new benchmark for evaluating AI agents on AWS infrastructure.","source":"hackernews_ai","url":"https://aws.amazon.com/about-aws/whats-new/2026/07/aws-bench/","published":"Sat, 25 Jul 2026 04:42:35 +0000"}]}]}