Rethinking On-Policy Distillation of Large Language Models II: One Training Example Paper • 2609.04172 • Published 13 days ago • 97
StudyBench: Can Self-Evolution Squeeze Textbooks for Olympiad Capability? Paper • 2609.00787 • Published 15 days ago • 21
PACE-Bench: Benchmarking Physics Adaptation via Code Evolution in Dynamic Environments Paper • 2608.14441 • Published Aug 14 • 29
Running 66 Don't Train the Model, Evolve the Harness 🌿 66 Evolving an agent's harness, not its model, on Harvey's LAB
Mutual Enhancement of Large and Small Language Models with Cross-Silo Knowledge Transfer Paper • 2312.05842 • Published Dec 10, 2023 • 1
ConCISE: Confidence-guided Compression in Step-by-step Efficient Reasoning Paper • 2505.04881 • Published May 8, 2025 • 1
Rethinking the Role of Efficient Attention in Hybrid Architectures Paper • 2606.15378 • Published Jun 13 • 20
δ-mem: Efficient Online Memory for Large Language Models Paper • 2605.12357 • Published May 12 • 133
Rethinking the Role of Efficient Attention in Hybrid Architectures Paper • 2606.15378 • Published Jun 13 • 20
ConCISE: Confidence-guided Compression in Step-by-step Efficient Reasoning Paper • 2505.04881 • Published May 8, 2025 • 1