Tech Meridian ← ENTITY INDEX
PROMY MERIDIAN RU

TOPIC · ENTITY #10764

contrastive reinforcement learning

Related event timeline, sources and context from the news index.

EVENT TIMELINE

1

RESEARCH · 1 SOURCE · arXiv cs.AI

ChronoSRL — temporal-geometry critic for self-supervised reinforcement learning (arXiv v1)

ChronoSRL is a self-supervised RL method that trains critic embeddings so distances correspond to goal-reaching time: reached goals are embedded at their travel time, unreached or off-trajectory goals are pushed beyond a discount horizon. It also predicts the distribution of goal-reaching times and time spent near the goal, and the policy is trained to prefer actions that reach goals sooner and more reliably; ChronoSRL outperforms contrastive, action-chunked contrastive, and survival-RL baselines on seven locomotion and navigation benchmarks and in simulated-to-real quadruped tasks (velocity tracking, goal-position reaching, box climbing).

6.0