NeoHorse-1: Towards Recursive Self-Improvement via Agentic Post-Training with Routing Harness
First concrete RSI mechanism proposed: a routing harness for agentic post-training.
3 items · 3 sources · 3 days
Operational story trace
Follow in this browser to see new updates on your Live feed.
Latest change
The newest paper (Sep 22) applies RSI to AI research agents automating their own R&D — training efficiency and inference optimization — rather than a fixed task domain.
Three papers in two weeks each propose a concrete recursive self-improvement (RSI) mechanism: a system that observes its own performance and converts that evidence into its next round of training. The pattern started with a general agentic post-training framework, then moved into domain-specific self-evolution loops for medical agents and AI research agents.
Arc
First concrete RSI mechanism proposed: a routing harness for agentic post-training.
Extends RSI to medical agents via clinically aligned self-evolution.
Extends RSI to AI research agents automating their own R&D pipeline.
First concrete RSI mechanism proposed: a routing harness for agentic post-training.
Extends RSI to medical agents via clinically aligned self-evolution.
Extends RSI to AI research agents automating their own R&D pipeline.
What to watch — open questions
Storylines are threaded mechanically from the feed: stories that share a distinctive anchor across multiple days and sources. Each item links to its original source. The evidence trace, current state, and open questions are written by the editor routine and refreshed whenever a new beat lands.