Story
arxiv_cs_ai ยท May 5, 2026 ยท paper
arxiv.orgMay 5, 2026
original source linked
In brief
Clinical LLMs are often scaled by increasing model size, context length, retrieval complexity, or inference-time compute, with the implicit expectation that higher accuracy implies safer behavior. This assumption is i...
Continue reading