Story

arxiv_cs_lg ยท May 1, 2026 ยท paper

Source brief

SAVGO: Learning State-Action Value Geometry with Cosine Similarity for Continuous Control

arxiv.orgMay 1, 2026
original source linked

In brief

While representation and similarity learning have improved the sample efficiency of Reinforcement Learning (RL), they are rarely used to shape policy updates directly in the action space. To bridge this gap, a geometr...

Continue reading

Read the original at arxiv.org โ†’Open in live feed

Earlier in this thread 4 items