EnvHarness: Awakening Static Worlds for Agent Learning Paper • 2608.19880 • Published 9 days ago • 265
SPADE: Self-Play in Adaptive Synthetic Executable Environments Paper • 2608.19197 • Published 10 days ago • 51 • 2
SPADE: Self-Play in Adaptive Synthetic Executable Environments Paper • 2608.19197 • Published 10 days ago • 51
SPADE Collection The full SPADE release: paper, model checkpoints, grounding corpora, and synthetic environments. • 4 items • Updated 9 days ago • 1
SPADE: Self-Play in Adaptive Synthetic Executable Environments Paper • 2608.19197 • Published 10 days ago • 51
From RLVR to RLSVR: Task Transformation Induces Self-Verifiable Rewards for Open-Ended LLM Self-Improvement Paper • 2607.23802 • Published Jul 26 • 106
From RLVR to RLSVR: Task Transformation Induces Self-Verifiable Rewards for Open-Ended LLM Self-Improvement Paper • 2607.23802 • Published Jul 26 • 106