{"date":"2026-07-26","title":"What happened in AI — Jul 26, 2026","generated_at":"2026-07-26T21:30:00Z","intro":["Two Show HN launches target coding agents' weakest spots today: architecture drift and lost session memory. A third open-source project, OpenLake, claims a 50% cut in long-horizon inference costs by offloading KV caches from GPU memory to RAM/NVMe, while a separate hypervisor project chases the same cost problem from the consumer-compute angle. Simon Willison surfaced the flip side of that cost pressure: an investigation into the relay market reselling pooled LLM API keys, some of it tied to fraud.","The day's bigger industry story is China's open-weight AI push hitting turbulence at the same time it gains ground. Moonshot AI shipped Kimi K3 and a widely syndicated wire report tracked growing US enterprise interest in cheaper Chinese models, even as DeepSeek paused a new funding round days after founder Liang Wenfeng's leaked remarks went viral."],"highlights":["OpenLake, a new open-source KV-cache offload engine, claims a 50% cut in long-horizon inference costs by moving caches from GPU memory to shared RAM/NVMe.","Boffin and CMEM target two coding-agent pain points from opposite ends: injecting architectural constraints per edit, and giving agents persistent memory across sessions.","Simon Willison surfaced an investigation into the relay market reselling pooled LLM API keys, some of it tied to fraud.","Moonshot AI launched Kimi K3, drawing renewed US commentary on the pace of Chinese open-weight releases.","DeepSeek paused a new fundraising round days after founder Liang Wenfeng's leaked remarks went viral, per South China Morning Post and The Times of India.","A widely syndicated wire report tracked growing US enterprise interest in cheaper, open Chinese models."],"article_count":13,"categories":[{"name":"Builders Target Coding Agents' Weak Spots: Architecture Drift and Lost Memory","slug":"coding-agents-architecture-drift-lost-memory","summary":"Two new tools attack coding-agent reliability from different angles: Boffin injects per-edit architectural constraints, while CMEM gives agents memory that survives across sessions instead of resetting each run.","articles":[{"title":"Show HN: Boffin – Staff-engineer layer for AI coding agents","summary":"Boffin adds a 'staff engineer' layer that routes per-edit architectural constraints into AI coding agents' output before it lands.","source":"hackernews_ai","url":"https://github.com/MicSm/boffin","published":"Sun, 26 Jul 2026 17:28:03 +0000"},{"title":"Show HN: CMEM – Persistent Memory for AI Coding Agents","summary":"CMEM gives coding agents persistent memory across sessions, targeting the context-loss problem that resets an agent's progress between runs.","source":"hackernews_ai","url":"https://cmem.ai","published":"Sun, 26 Jul 2026 14:54:58 +0000"}]},{"name":"Inference Gets Cheaper on the Legit Side, Murkier on the Black Market","slug":"inference-cost-legit-black-market","summary":"Two open-source projects chase cheaper inference through KV-cache offloading and consumer-compute hosting, while an investigative report maps a parallel relay market reselling pooled API tokens at a discount.","articles":[{"title":"Show HN: Cuts Long Horizon Inference Costs by 50% via external KV Cache Offload","summary":"OpenLake is an open-source storage engine that offloads LLM KV caches from GPU memory to a shared tier of RAM and NVMe, cutting long-horizon inference costs by half.","source":"hackernews_ai","url":"https://github.com/openlake-project/openlake","published":"Sun, 26 Jul 2026 13:02:21 +0000"},{"title":"Show HN: I built a hypervisor and client for inference on consumer compute","summary":"Scalattice built a hypervisor for running inference workloads on consumer-grade compute, aiming to make idle consumer hardware usable for LLM serving.","source":"hackernews_ai","url":"https://scalattice.com/blog/openai-sdk-scalattice/","published":"Sun, 26 Jul 2026 05:43:09 +0000"},{"title":"An Inside Look at the Relay Market Powering Token Resellers and Fraud","summary":"Investigative reporting picked up by Simon Willison maps a relay market where pooled API keys resell discounted LLM tokens, some of it tied to fraud.","source":"simon_willison","url":"https://simonwillison.net/2026/Jul/26/relay-market/#atom-everything","published":"2026-07-26T19:30:54+00:00"}]},{"name":"China's Open-Weight Wave Keeps Coming, Even as DeepSeek Hits Turbulence","slug":"china-open-weight-deepseek-turbulence","summary":"Moonshot shipped Kimi K3 and a widely syndicated wire report tracked rising US enterprise interest in cheaper Chinese models, while DeepSeek paused a new funding round days after founder Liang Wenfeng's leaked remarks went viral.","articles":[{"title":"Moonshot AI launches Kimi K3 and sparks US AI concerns","summary":"Moonshot AI released Kimi K3, drawing renewed US commentary on the pace of Chinese open-weight model releases.","source":"search_cn_open_weight_labs","url":"https://news.google.com/rss/articles/CBMiakFVX3lxTE05QmpoTTNLTGNud2dVeTZIWldmRW1hYlFDNEgtejdyMDhCUzFNN3VXajVjTEJ1T3lyQ2w2Rkw1Q0dlLWwzTFBkRGt4YWRGOTFtUkxtc3daWUlycHdVNkh0UkZSem1MRHJPWnc?oc=5","published":"Sun, 26 Jul 2026 02:17:59 GMT"},{"title":"China's DeepSeek pauses new fundraising days after founder Liang Wenfeng's AI remarks go viral - The Times of India","summary":"DeepSeek paused a new funding round in the days following founder Liang Wenfeng's viral remarks, per The Times of India.","source":"search_cn_open_weight_labs","url":"https://news.google.com/rss/articles/CBMihgJBVV95cUxNenNNam5MV0lndndMUEFPQU5hMEMxWVo2bGM2QnV1MzlZcHJZbkg5am40T3dxeUdVQ1gybVByaWJ5TnRCRjdZdkdhbGo2UXhaNzVqWV9xR3RIZC0wYTRpMHF4WFNjVjdSN0FJRnBMeng0T3oxdHpSMW9qd1J2WExES0xnWklybFVFNTdDVEpzamE5WlFVNnIyRzBkMkhCSUFQOUxJSXpLaXpTS3h1azBEalB5R0pTSkJWaHdsakJMaWhWYXN5R2lJVHRWd0p1ZXVRQjdlc244ZVdWWHo4R3kxcWNFOHM2Qlh1Yk5mcmV1b2JMZ2xNQ1Y3N21ESjVUVVBCTzFUWHBR0gGLAkFVX3lxTE81dFZFQkhiV05Wal9iU2M0QURRMTl3cC1YcnBlTElGYjJXZTEyU2FueVBZYmZ6TTJPZTJNay1kS04yY0FGdUUzNmU4VEJhc2RkU09LMnFIc3VlLW9wdHpRbnZEdHJqNDdvbVVMMTdaUG5YWlJTbVh4Y1pna2pjQnVrZ05zTG9BNW5kRFFsdHRGcG5FNVdlS21UUXZzMUNLNXZ3cHdtcENVUk1YeUZGd2JZTjNkX1haZDRXZTBRNUxGWTl3MEhtd0Z5R056emNSZXN4V3luYndMcEtIQ0o5NktiOTYyRElTUGFIbFZlY2VLV0JHN2p2QmhwRFlNUXozMFpjM1JQcUZzMjBkOA?oc=5","published":"Sun, 26 Jul 2026 09:52:00 GMT"},{"title":"DeepSeek founder Liang Wenfeng's surprising philosophy revealed in leak - South China Morning Post","summary":"Leaked remarks attributed to DeepSeek founder Liang Wenfeng went viral, per South China Morning Post, without official confirmation from the company.","source":"search_cn_open_weight_labs","url":"https://news.google.com/rss/articles/CBMi1wFBVV95cUxPUjhianFTdVFTZEdSeE1wcS1BbURJekQ1SllTMzluSGhhaTQ2eWlOZk5hemxoVDNualZhTWxVaXhUUXZUZlBnY2hJUFRoMFBPZlV2RHhTUzk4LTBKSFR1SXdVaW5uejdnaGJaNzVfeFZwb1p2UmpxN2FpVXJnWEl0MzYxYlJPbDlHWnRBdTA0X0Q2XzFzeEV2VTUza3JVNGNLaUhOVHFGeW1heGR3NVRpMGVqLWd0bE83S2hvNFc0Ynp4Tlptamc4ckgwQ09CTkk1QUFDSnJzUdIB1wFBVV95cUxQS3RndlZuRWxVUEh6UU1PSlpzdzRXMUFydE9uNnBSQmg0VlN2VXd2ZXpYZ19ITDJIN2drbGFnRDFvMXZ6UW1nTDhkWmtnZEkwc1BPRW9oTnExQXFHdkNYQkJMZGh2NWhNNGhrSFVHZDVWNnJmWURZRkNsUV91dXd1MDM2QTFQb2VyeV9XYXBnZ0MyX3FMSmszWDFPRzFxY0hGUXItRGxUVzN6dmRubk4zUEZ6MC1WMUhIbjFya01QQTV6enJ3clpaM2xEZjZaTTRuZndoWF9jQQ?oc=5","published":"Sun, 26 Jul 2026 12:00:08 GMT"},{"title":"Cheaper, open and intelligent: Chinese AI models gain ground, as they make inroads in the US - The Washington Post","summary":"A widely syndicated wire report tracks growing US enterprise interest in cheaper, open Chinese models even as export controls remain in place.","source":"search_cn_open_weight_labs","url":"https://news.google.com/rss/articles/CBMiygFBVV95cUxNS3lwcWJjLVNidGk2RzR4Wm93am0ycS1fbmNyZkFGcEdpM1l0blpxTzNhYVlIanRZaWViRGI3d0V3VE4tdHpaWDZzZC1vNFF0clBYZ01GYXFfRUx4Z2Jsc2RjNkN2RnpNQjZhb2ZpS1NEMklpanl3aVBlY0VjRENmLVk4M0lVYWpMZUNRYU9FZUtlNVFTeXlPcWxoSGtHVmNqejhWcFk1ZU54WWl4b25BYnE4eC1sSXo2M2xyXy1aZDBIMktWRmJ3ZlRR?oc=5","published":"Sun, 26 Jul 2026 04:35:38 GMT"}]}]}