Story
arxiv_cs_cl ยท Oct 7, 2026 ยท paper
Source brief
From Pareto to Preference: Personalized Test-Time Scaling via Amortized Agentic Policy Discovery
arxiv.orgOct 7, 2026
original source linked
In brief
Test-time scaling (TTS) improves the reasoning capabilities of large language models by allocating additional inference computation. Existing approaches to improving TTS efficiency largely optimize accuracy against on...
Feed lens
agenticevaluation