EnvHarness: Awakening Static Worlds for Agent Learning Paper • 2608.19880 • Published 16 days ago • 274
SPADE: Self-Play in Adaptive Synthetic Executable Environments Paper • 2608.19197 • Published 17 days ago • 52 • 2
SPADE: Self-Play in Adaptive Synthetic Executable Environments Paper • 2608.19197 • Published 17 days ago • 52
SPADE Collection The full SPADE release: paper, model checkpoints, grounding corpora, and synthetic environments. • 4 items • Updated 16 days ago • 1
SPADE: Self-Play in Adaptive Synthetic Executable Environments Paper • 2608.19197 • Published 17 days ago • 52
From RLVR to RLSVR: Task Transformation Induces Self-Verifiable Rewards for Open-Ended LLM Self-Improvement Paper • 2607.23802 • Published Jul 26 • 106
From RLVR to RLSVR: Task Transformation Induces Self-Verifiable Rewards for Open-Ended LLM Self-Improvement Paper • 2607.23802 • Published Jul 26 • 106