{"date":"2026-08-29","title":"What happened in AI — Aug 29, 2026","generated_at":"2026-08-29T21:11:53Z","intro":["OpenAI is cutting Cursor off from its models on November 12, following SpaceX's acquisition of the coding-agent startup — the clearest sign yet that the Musk-Altman rivalry now shapes which AI tools engineers can use, even though OpenAI's models make up only 5% of Cursor's traffic today.","Elsewhere, agent tooling matured on its own terms: a new git-worktree orchestrator and a benchmark-backed multi-agent coding harness both tackle running several coding agents on one repo safely, and a published governance protocol showed an autonomous research agent correctly refusing to certify its own flawed results."],"highlights":["OpenAI cuts Cursor off from its models on Nov 12, following SpaceX's acquisition of the startup — OpenAI models are only 5% of Cursor's traffic today.","Metis, a 5-role recursive coding-agent harness, scored 82% vs. OpenCode's 67% on Terminal-Bench 2.1 using the same DeepSeek V4 Flash model for both.","LaneGate locks the files each task touches across separate git worktrees, so parallel coding agents can't silently overwrite each other's edits.","A \"governed-pass\" protocol shows an autonomous Claude research agent, after 4+ days of work, correctly refusing to certify its own results — a flaw two independent reviewers built from the same spec had both missed.","TOTVS builds domain-specific MCP tools instead of generic query-generation ones for agent data access, and says an RDF/OWL semantic layer lifted LLM response precision by about 40%.","FreeToken (UC Berkeley/MIT) runs frontier MoE models like DeepSeek-V4-Flash and GLM-5.2 on consumer GPUs, claiming 3-4x faster decode than Ollama or llama.cpp."],"article_count":9,"categories":[{"name":"Coding-Agent Tooling & Access","slug":"coding-agent-tooling-access","summary":"The tools agents run on kept advancing even as OpenAI moved to cut off the one connecting Cursor to its models — a reminder that the agent-tooling supply chain is now geopolitical as much as technical.","articles":[{"title":"[AINews] OpenAI shuts off Cursor","summary":"OpenAI will end Cursor's model access on Nov 12 following SpaceX's acquisition of the startup, citing Musk-linked companies \"violating contracts\" — though OpenAI's own models are only 5% of Cursor's traffic today.","source":"latent_space","url":"https://www.latent.space/p/ainews-openai-shuts-off-cursor","published":"Sat, 29 Aug 2026 05:11:52 GMT"},{"title":"Show HN: Metis – An agent harness pushing DeepSeek to Opus-tier coding (82%)","summary":"This open-source, five-role recursive coding harness (Coordinator, Planner, Implementer, Reviewer, Verifier) scored 82% vs. OpenCode's 67% on Terminal-Bench 2.1, using the identical DeepSeek V4 Flash model for both.","source":"hackernews_ai","url":"https://github.com/Wholiver/metis","published":"Sat, 29 Aug 2026 02:36:07 +0000"},{"title":"LaneGate – Git-native worktree orchestrator for AI agents","summary":"This orchestrator gives each coding-agent task its own git worktree and locks the files it declares it will touch, preventing parallel agents from colliding on the same repo.","source":"hackernews_ai","url":"https://github.com/sudheerdvn/lanegate","published":"Sat, 29 Aug 2026 05:36:23 +0000"},{"title":"Show HN: DeepSeekGUI – A Windows desktop client for DeepSeek's coding agent","summary":"A new Electron client wraps DeepSeek's open-source Harness coding agent for Windows, adding an installer, system tray, and built-in browser on top of the official web UI.","source":"hackernews_ai","url":"https://github.com/See-Sol-Lab/DeepSeekGUI","published":"Sat, 29 Aug 2026 11:02:05 +0000"}]},{"name":"Agent Data & Governance","slug":"agent-data-governance","summary":"Two pieces looked at trust from the infrastructure level up: how agents access enterprise data, and how they judge their own work.","articles":[{"title":"Architecting the Data Layer for AI Agents: From Transactional Systems to MCP and Semantic Models","summary":"TOTVS builds domain-specific MCP tools instead of generic query-generation tools to avoid prompt injection, and reports its RDF/OWL semantic layer lifted LLM response precision by about 40%.","source":"infoq_ai_ml","url":"https://www.infoq.com/presentations/enterprise-data-architecture-ai-agents/?utm_campaign=infoq_content&utm_source=infoq&utm_medium=feed&utm_term=AI%2C+ML+%26+Data+Engineering","published":"Sat, 29 Aug 2026 11:00:00 GMT"},{"title":"One prompt, five days, and an AI agent that refused to certify itself","summary":"After running for over four days, an autonomous Claude research agent under this governance protocol correctly refused to certify its own results and halted — a flaw that two independent reviewers built from the same spec had both missed.","source":"hackernews_ai","url":"https://github.com/Framework-Drift/governed-pass","published":"Sat, 29 Aug 2026 01:16:04 +0000"}]},{"name":"Models & Efficient Inference","slug":"models-efficient-inference","summary":"Chinese labs kept racing on price and benchmarks while a new open-source engine made frontier-scale MoE models runnable on ordinary consumer hardware.","articles":[{"title":"Qwen 3.8 Flash Reduces Costs to One-Third of DeepSeek-V4-Flash","summary":"Alibaba's Qwen 3.8 Flash reportedly cuts inference costs to one-third of DeepSeek-V4-Flash's, per KuCoin.","source":"search_cn_open_weight_labs","publisher_name":"KuCoin","publisher_domain":"kucoin.com","url":"https://news.google.com/rss/articles/CBMimAFBVV95cUxQeE50b2tjajhZd0N6a0lvX3VuVU4xckJWMXRBTkJZOWotMVdSMXVaWThGOTVxMHlBNzA3NkxEekFwaTVyeW1GU3pERHZQa0pObHdSNTdVNHQ1V0s0NTEweDJ1UUFfdHJDWG5ER0xnV1dEM0J3c093RUZXUldpNzJSNXZxc0tRQkFNaGI4MlNLSkU3VXpXdHlfOA?oc=5","published":"Sat, 29 Aug 2026 05:15:10 GMT"},{"title":"Tencent unveils AI model it says outperforms Z.ai, Moonshot","summary":"Tencent unveiled a new model it claims outperforms rivals Z.ai and Moonshot, per Tech in Asia.","source":"search_cn_open_weight_labs","publisher_name":"Tech in Asia","publisher_domain":"techinasia.com","url":"https://news.google.com/rss/articles/CBMiiAFBVV95cUxQN3JZTVI0akE0dVk4NmtyN0ZjQlVGcFdZWGY4U1Zsa3VZMXVaNkJudGNoVDZaRzhyMEoxbkNQVlhkaVdJVXo3TkhRblJxVVRCRGNzYWFjaDk5MEVMNWF5dlhZWENqOU41Vnh1TE9IYnZwYm1MczJqZlV5Zk13aHNodlFLSTVwUFRz?oc=5","published":"Sat, 29 Aug 2026 03:11:00 GMT"},{"title":"FreeToken Unlocks Frontier MoE Inference on Consumer Hardware via Dynamic Co-Execution","summary":"This open-source inference engine from UC Berkeley and MIT splits MoE token computation between CPU and GPU in real time, running models like DeepSeek-V4-Flash and GLM-5.2 on single consumer or workstation GPUs with claimed 3-4x faster decode than Ollama or llama.cpp.","source":"infoq_ai_ml","url":"https://www.infoq.com/news/2026/08/freetoken-local-inference/?utm_campaign=infoq_content&utm_source=infoq&utm_medium=feed&utm_term=AI%2C+ML+%26+Data+Engineering","published":"Sat, 29 Aug 2026 05:05:00 GMT"}]}]}