Dream-RSI: Recursive Self-Improvement through Evolving Worlds Paper • 2609.14858 • Published 5 days ago • 225
Benchmark Radar: A Living Database and Search Engine for AI Benchmarks and Evaluation Paper • 2609.11115 • Published 9 days ago • 164
Recursive Code World Models: Building Complex Worlds through Recursive Scene Programs Paper • 2609.11499 • Published 9 days ago • 33
FlowBalance: Verifier-Grounded Self-Improvement from On-Policy Reasoning Experience Paper • 2609.03241 • Published 16 days ago • 53
RISE: Recursive Improvement via Self-Extrapolating Policy Distillation Paper • 2609.05295 • Published 15 days ago • 17
LatentPress: Context Compression Beyond Text and Vision Paper • 2609.01507 • Published 18 days ago • 118
Random Attention: Rethinking KV Cache Eviction for Efficient Reasoning Paper • 2609.03430 • Published 16 days ago • 186
Influence-Directed Distillation: Solving the Diversity Bottleneck in Sampled-Token On-Policy Distillation Paper • 2608.29846 • Published 20 days ago • 15
It Takes Two to Match: Co-Evolving Generative Retriever with Reinforcement Learning Paper • 2609.00638 • Published 18 days ago • 65
Does On-Policy Distillation Really Distill? From Noisy Teacher to Self-Improvement Paper • 2608.31046 • Published 19 days ago • 95
J-Zero: Unified Challenger--Solver--Judge Co-Evolution from Zero Data Paper • 2608.26582 • Published 23 days ago • 41
EnvHarness: Awakening Static Worlds for Agent Learning Paper • 2608.19880 • Published 30 days ago • 174
SPADE: Self-Play in Adaptive Synthetic Executable Environments Paper • 2608.19197 • Published about 1 month ago • 54
SimpleOPD: Simple Tokenizer-Agnostic On-Policy Distillation for Long-Context Reasoning Paper • 2608.14277 • Published Aug 14 • 36