{"date":"2026-09-18","title":"What happened in AI — Sep 18, 2026","generated_at":"2026-09-18T21:14:17Z","intro":["Coding-agent tooling kept consolidating and scaling up today: Claude Code added AGENTS.md as a fallback convention, and production teams showed agents doing large operational work rather than just writing code — DoorDash's multi-agent system retired 60,000 stale feature flags across 623 repos, and Google detailed using agentic AI to help secure hundreds of millions of lines of its infrastructure code.","Evaluation is getting institutional backing to match: Anthropic and Accenture committed over $1 billion combined to build independent evaluation capacity, and OpenAI published an internal triage framework for reporting model misalignment. Chinese labs kept shipping — Zhipu's GLM-5.3-FlashX and Moonshot's climb toward a $50 billion valuation — but a new report says their revenue still trails OpenAI and Anthropic despite lower costs."],"highlights":["Claude Code (v2.1.277) now falls back to AGENTS.md when no CLAUDE.md is present, folding a rival convention into its own config discovery.","DoorDash's multi-agent LLM system retired 60,000 stale feature flags across 623 repositories, with engineer approval gates in the loop.","Anthropic and Accenture will invest more than $1 billion combined to build independent evaluation capacity for frontier AI.","Zhipu's GLM-5.3-FlashX runs near 200 tokens/sec on roughly 100,000 domestically made accelerators.","A new report says OpenAI and Anthropic's Chinese rivals still capture only a fraction of their revenue despite lower operating costs."],"article_count":16,"categories":[{"name":"Coding-Agent Standards Consolidate Around AGENTS.md and MCP","slug":"coding-agent-standards-consolidate","summary":"Three signals point the same direction: coding-agent tooling is consolidating around shared conventions instead of one-off formats — Claude Code adopting AGENTS.md, a dedicated fast-decision model slotting into the agent loop, and public debate over whether Skills has already superseded MCP and RAG.","articles":[{"title":"Claude Code Adds AGENTS.md Support as a CLAUDE.md Fallback","summary":"Claude Code v2.1.277 now falls back to AGENTS.md when no CLAUDE.md is present, folding a competing convention into its own config discovery.","source":"simon_willison","url":"https://simonwillison.net/2026/Sep/18/thariq-shihipar/","published":"2026-09-18T19:09:27+00:00"},{"title":"Should You Read the Code, Is RAG Dead, and Did Skills Kill MCP?","summary":"GitHub's podcast crew debate whether Anthropic's Skills format has displaced MCP and whether RAG still earns its keep against long-context coding agents.","source":"github_blog_ai_ml","url":"https://github.blog/ai-and-ml/should-you-read-the-code-is-rag-dead-and-did-skills-kill-mcp/","published":"2026-09-18T15:00:00Z"},{"title":"What Is Jev? TypeSafe AI's Fast \"System One\" Model for Agent Loops","summary":"TypeSafe AI's Jev is a small \"System One\" model built for fast, structured decisions inside a LangChain agent loop, distinct from the LLM doing the slow reasoning.","source":"langchain_blog","url":"https://www.langchain.com/blog/building-a-harness-with-jev","published":"2026-09-18T01:47:48Z"}]},{"name":"Agents Move From Coding Assistants to Large-Scale Ops Automation","slug":"agents-move-to-ops-automation","summary":"Today's deployment stories move past coding assistance into large-scale operational work: DoorDash's agents retired 60,000 feature flags, Google secured hundreds of millions of lines of infrastructure code, and AWS packaged agent skills to handle model deployment end to end.","articles":[{"title":"DoorDash Uses Multi-Agent LLMs to Clean Up 60,000 Feature Flags","summary":"DoorDash's multi-agent LLM system retired 60,000 stale feature flags across 623 repositories, combining live experiment data via MCP with engineer approval gates.","source":"infoq_ai_ml","url":"https://www.infoq.com/news/2026/09/doordash-feature-flag-cleanup/?utm_campaign=infoq_content&utm_source=infoq&utm_medium=feed&utm_term=AI%2C+ML+%26+Data+Engineering","published":"2026-09-18T13:50:00Z"},{"title":"How Google Uses Agentic AI to Secure Hundreds of Millions of Lines of Code","summary":"Google detailed how it uses agentic AI to scan and secure hundreds of millions of lines of infrastructure code against emerging AI-assisted exploit techniques.","source":"google_cloud_blog","url":"https://cloud.google.com/blog/topics/systems/using-ai-agents-to-secure-google-infrastructure/","published":"2026-09-18T16:00:00Z"},{"title":"AWS Ships Six Coding-Agent Skills to Deploy Hugging Face Models on SageMaker","summary":"AWS packaged six open-source agent skills that let a coding agent deploy Hugging Face models on SageMaker AI, auto-selecting serving containers and autoscaling.","source":"aws_ml_blog","url":"https://aws.amazon.com/blogs/machine-learning/deploy-hugging-face-models-on-amazon-sagemaker-ai-with-coding-agents/","published":"2026-09-18T15:25:23Z"}]},{"name":"AI Infrastructure and Inference Scaling","slug":"ai-infrastructure-and-inference-scaling","summary":"AWS's new GPU-aware inference router and vLLM's use of NVIDIA's hardware video decoders both squeeze more inference throughput out of existing GPUs rather than just adding more of them.","articles":[{"title":"Amazon SageMaker HyperPod Inference Gateway Routes by Live GPU Load","summary":"SageMaker HyperPod's new Inference Gateway is a Kubernetes-native, GPU-aware router on EKS that uses live GPU signals to cut first-token latency.","source":"aws_ml_blog","url":"https://aws.amazon.com/blogs/machine-learning/introducing-amazon-sagemaker-hyperpod-inference-gateway/","published":"2026-09-18T13:08:34Z"},{"title":"vLLM Scales Multi-GPU Video Captioning With NVIDIA's Video Decoders","summary":"vLLM detailed using NVIDIA's hardware video decoders to scale multi-GPU video captioning and description workloads.","source":"vllm_blog","url":"https://vllm.ai/blog/2026-09-18-pynvvideocodec","published":"2026-09-18T00:00:00Z"}]},{"name":"Open-Weight Releases From Chinese Labs","slug":"open-weight-releases-from-chinese-labs","summary":"Zhipu pushed a fast open-weight model on domestic accelerators while DeepSeek's newly open-sourced execution harness reframes what counts as \"industrial-grade\" agent engineering.","articles":[{"title":"Zhipu's GLM-5.3-FlashX Hits Near 200 Tokens/s on Domestic Chips","summary":"Zhipu's GLM-5.3-FlashX runs near 200 tokens/sec on a cluster of roughly 100,000 domestically made accelerators, its latest bet on chip self-sufficiency.","source":"search_cn_open_weight_labs","publisher_name":"Pandaily","publisher_domain":"pandaily.com","url":"https://news.google.com/rss/articles/CBMifEFVX3lxTE9qcmhXanNqRDFRWTZFUlc5S1VwYnhQWlZtLWI1eWROOTlSdU1xQTJJVXlzaG9mdVlQRGZHUEpFVWpUZmxKdVptX2J2QUxxeUQ0R1JnR0pENmFsbGRBczFhcGV2Njc0T3k0M1hkVzcxamx5S05zWVFmS0J4cy0?oc=5","published":"2026-09-18T07:54:43Z"},{"title":"DeepSeek's Harness Goes Open Source, Sparking Debate on Agent-Engineering Boundaries","summary":"36Kr examines DeepSeek's newly open-sourced execution harness and the industrial-grade boundaries it sets for production agent engineering.","source":"search_cn_open_weight_labs","publisher_name":"36 Kr","publisher_domain":"eu.36kr.com","url":"https://news.google.com/rss/articles/CBMiU0FVX3lxTE9FWmh2RlZyaENsWUhseTFLOGVZakg4SGMzeWxTQm1WcnNZVWRZdTBuSUdJWm8xYTVYQk92UnNLYUdLdVd2N2pzU0J6V2E5THlFZmpR?oc=5","published":"2026-09-18T09:34:38Z"}]},{"name":"Evals and Safety Get Institutional Backing","slug":"evals-and-safety-get-institutional-backing","summary":"Independent evaluation of frontier AI is getting real money and process behind it — Anthropic and Accenture's $1 billion-plus commitment, OpenAI's internal misalignment-reporting framework, and a new open benchmark for agent memory.","articles":[{"title":"Anthropic and Accenture Commit Over $1 Billion to Independent AI Evaluation","summary":"Anthropic and Accenture will invest more than $1 billion combined to build independent, embedded evaluation capacity for frontier AI models.","source":"anthropic_newsroom","url":"https://www.anthropic.com/news/accenture-embedded-evaluation","published":"2026-09-18T00:00:00+00:00"},{"title":"OpenAI Introduces a Triage Framework for Reporting Model Misalignment","summary":"OpenAI published a triage framework letting employees flag suspected model misalignment, plus initial case studies of unexpected model behavior.","source":"infoq_ai_ml","url":"https://www.infoq.com/news/2026/09/openai-misalignment-framework/?utm_campaign=infoq_content&utm_source=infoq&utm_medium=feed&utm_term=AI%2C+ML+%26+Data+Engineering","published":"2026-09-18T05:05:00Z"},{"title":"An Open Benchmark Aims to Standardize How Agent Memory Systems Are Compared","summary":"A new open leaderboard standardizes Add/Search benchmarking for agent memory systems, aiming to stop each vendor from grading its own homework.","source":"hackernews_ai","url":"https://agentmemoryleaderboard.ai/","published":"2026-09-18T02:52:01Z"}]},{"name":"Chinese AI Labs' Revenue Gap Widens as Compute Scales","slug":"chinese-ai-labs-revenue-gap-widens","summary":"Even as Moonshot's valuation and Z.ai's compute position keep climbing, a fresh report says Chinese model providers are still capturing only a fraction of OpenAI and Anthropic's revenue.","articles":[{"title":"OpenAI and Anthropic's Chinese Rivals Face a Revenue Reality Check","summary":"A new report says OpenAI and Anthropic's Chinese rivals are still capturing only a fraction of their revenue despite running at lower cost.","source":"search_cn_open_weight_labs","publisher_name":"tradingview.com","publisher_domain":"tradingview.com","url":"https://news.google.com/rss/articles/CBMixAFBVV95cUxQdG1ReG9VNFlaandoNkp0R0hUcE90VXMySWROTWl4a2xMRk5HR1Uyc01lUGlEeExOUG12SjdkVTZqY1hldThHMXNWQ3JVRzZYY3Q4OTlHWjFhak9rQ1lJcUdpVXlWVHFmTXdHbWl6YU5BQ0lvRHpNVVVOcnJhWEJEQ1lrbXdsQXk1RnFKREFnZkhqdFYwdm9YX0Vna0V2SEFZMlljVTMyVVN5TmJjNTc5UndVUWNwWW1kdXJBclBQZTU4VmFC?oc=5","published":"2026-09-18T19:33:17Z"},{"title":"Inside the $50 Billion Rise of Moonshot AI","summary":"ThinkChina traces Moonshot AI's climb to a $50 billion valuation and the bets behind it.","source":"search_cn_open_weight_labs","publisher_name":"ThinkChina","publisher_domain":"thinkchina.sg","url":"https://news.google.com/rss/articles/CBMifkFVX3lxTE5ETnhGRTBWZWc4VERQd2tIMC1oYXFRSWw1TTYyZHlTaXZaclZzLTNJVU84bnhmdm1KcHZsU0ZuSTIxOTBndlUwNWQxQ3JHRU82dlFndmoxbFl2b3lpekpaQ01DVWUxa2kyMDdRdjQ4SnNSY0V3ZkhRV3d3SGVGQQ?oc=5","published":"2026-09-18T03:03:31Z"},{"title":"Z.ai's Next Growth Engine Starts Where Compute Scarcity Ends","summary":"Digitimes argues Z.ai's next growth phase begins as domestic compute scarcity eases, shifting the constraint from chips to product.","source":"search_cn_open_weight_labs","publisher_name":"digitimes","publisher_domain":"digitimes.com","url":"https://news.google.com/rss/articles/CBMikwFBVV95cUxNeUVTUDc1bkM4U2x0ZmNoWGxLbUpDTk5xMndNY2FxWmt0WWdMMUkyMHRxWGlwUmxCajhHdjEzRGhCaVFBaUNCUjRFUXpjV1ZqMUNPUUk1Z2RWcldBVDhxZnM5SmpxMGZ3UVRIa3RtNWY4ZDVIUzFUcEZNaDNrUHhXXzdCNlIxaGtjT0xVLWFFcHVQNXM?oc=5","published":"2026-09-18T00:29:14Z"}]}]}