Dream-RSI: Recursive Self-Improvement through Evolving Worlds Paper • 2609.14858 • Published 6 days ago • 230
Benchmark Radar: A Living Database and Search Engine for AI Benchmarks and Evaluation Paper • 2609.11115 • Published 10 days ago • 167
Recursive Code World Models: Building Complex Worlds through Recursive Scene Programs Paper • 2609.11499 • Published 10 days ago • 33
FlowBalance: Verifier-Grounded Self-Improvement from On-Policy Reasoning Experience Paper • 2609.03241 • Published 17 days ago • 53
RISE: Recursive Improvement via Self-Extrapolating Policy Distillation Paper • 2609.05295 • Published 16 days ago • 17
LatentPress: Context Compression Beyond Text and Vision Paper • 2609.01507 • Published 19 days ago • 118
Random Attention: Rethinking KV Cache Eviction for Efficient Reasoning Paper • 2609.03430 • Published 17 days ago • 186
Random Attention: Rethinking KV Cache Eviction for Efficient Reasoning Paper • 2609.03430 • Published 17 days ago • 186
Influence-Directed Distillation: Solving the Diversity Bottleneck in Sampled-Token On-Policy Distillation Paper • 2608.29846 • Published 21 days ago • 15
It Takes Two to Match: Co-Evolving Generative Retriever with Reinforcement Learning Paper • 2609.00638 • Published 19 days ago • 65
Does On-Policy Distillation Really Distill? From Noisy Teacher to Self-Improvement Paper • 2608.31046 • Published 20 days ago • 96
J-Zero: Unified Challenger--Solver--Judge Co-Evolution from Zero Data Paper • 2608.26582 • Published 24 days ago • 41
EnvHarness: Awakening Static Worlds for Agent Learning Paper • 2608.19880 • Published about 1 month ago • 174
EnvHarness: Awakening Static Worlds for Agent Learning Paper • 2608.19880 • Published about 1 month ago • 174 • 2
EnvHarness: Awakening Static Worlds for Agent Learning Paper • 2608.19880 • Published about 1 month ago • 174
SPADE: Self-Play in Adaptive Synthetic Executable Environments Paper • 2608.19197 • Published Aug 19 • 54