Story

arxiv_cs_cl ยท Oct 7, 2026 ยท paper

Source brief

From Pareto to Preference: Personalized Test-Time Scaling via Amortized Agentic Policy Discovery

arxiv.orgOct 7, 2026
original source linked

In brief

Test-time scaling (TTS) improves the reasoning capabilities of large language models by allocating additional inference computation. Existing approaches to improving TTS efficiency largely optimize accuracy against on...

Feed lens
agenticevaluation

Continue reading

Read the original at arxiv.org โ†’Open in live feedRead that dayโ€™s brief

Earlier in this thread 4 items