As multimodal large language models (MLLMs) become more capable and widely deployed, concerns about privacy and safety have become increasingly pressing. Machine unlearning offers one approach to addressing these conc... Context & related coverage →
arxiv.org · 2026-10-07 · Ranked: agent + eval match · research watch · fresh 0.92 · score 2.38
Language-model agents are increasingly asked to carry out open-ended scientific research, yet their results are usually graded against a known answer, a rubric, or a language-model reviewer, none of which can tell whe... Context & related coverage →
A survey conducted by independent research firm Coleman Parkes on behalf of Undo, a company focused on scaling AI-powered root-cause analysis, found that while AI coding agents have accelerated code generation, they h... Context & related coverage →
During its recent “Birthday Week”, Cloudflare announced Clef, a set of open-weight AI models designed to choose between predefined options rather than generate text. Cloudflare released 9B- and 27B-parameter models, a... Context & related coverage →
As previously promised , here's Anthropic's new fast, low cost model: Introducing Claude Haiku 5.5 . The previous Haiku, 4.5, was very much showing its age. It came out almost a year ago , and was priced at $1/million... Context & related coverage →
arxiv.org · 2026-10-07 · Ranked: eval match · research watch · fresh 0.92 · score 2.23
Speculative decoding accelerates large language model inference by using a low-cost draft model to propose tokens that the full-size target model verifies in parallel. Parallel and semi-autoregressive (semi- AR) draft... Context & related coverage →
primeintellect.ai · 2026-10-08 · Ranked: community signal · fresh 0.97 · score 2.11 · Context
Kernels encode the inductive bias of a wide range of machine learning models, yet automated kernel design faces a fundamental dilemma. A fixed grammar of base kernels and operators guarantees validity but limits the s... Context & related coverage →
arxiv.org · 2026-10-07 · Ranked: evaluation match · research watch · fresh 0.92 · score 1.95
Hallucinated information can propagate through multi-stage LLM systems and become part of the context for subsequent reasoning. Existing studies of post-hallucination reasoning (PHR) mainly characterize changes in fin... Context & related coverage →
Managed Deep Agents is the simplest way to build, deploy, and run agents in production. The 0.8 release adds support for user-owned credentials, user-level memory, HTTP channels, file transfer in Slack and a pre-built... Context & related coverage →
LangChain introduces LangSmith Fine-Tuning and SmithTune, a CLI built for post-training models. Train specialized models without building data pipelines by hand. Context & related coverage →
Business Wire · 2026-10-07 · Ranked: agent match · community signal · fresh 0.82 · score 1.85 · Context