Preprint2025
Absolute Zero
Unknown
A reinforcement learning paradigm where a single model learns to propose tasks maximizing its own learning progress and improves reasoning by solving them, enabling continuous self-improvement without human data.
0May 1, 2025Reinforcement LearningReasoning