Story
arxiv_cs_lg ยท May 1, 2026 ยท paper
Source brief
SAVGO: Learning State-Action Value Geometry with Cosine Similarity for Continuous Control
arxiv.orgMay 1, 2026
original source linked
In brief
While representation and similarity learning have improved the sample efficiency of Reinforcement Learning (RL), they are rarely used to shape policy updates directly in the action space. To bridge this gap, a geometr...
Continue reading