Science or Slop?: Benchmarking and Mitigating Scientific Slop in AI-Generated Papers Paper • 2610.00531 • Published 12 days ago • 62
Transformers Stop Thinking Too Early, and a Tiny LoRA Fixes It Paper • 2609.36585 • Published 13 days ago • 71
OmniTaskonomy: When Does Visual Generation Improve Visual Understanding? Paper • 2609.38079 • Published 13 days ago • 57
Knowing When Thinking Is Not Enough: Teaching Small Reasoning Models to Reason Beyond Their Parametric Knowledge Paper • 2609.34327 • Published 14 days ago • 41
FuseReg: Regularizing Layer Fusion Mitigates the Reconstruction-Generation Gap in Representation Autoencoders Paper • 2609.31620 • Published 17 days ago • 145
MLLMs Hallucinate when Information Distribution Drifts in Synergy Heads Paper • 2609.09206 • Published Sep 5 • 11
ChronoVision: Temporal Reasoning via Latent State Reconstruction Paper • 2608.05631 • Published Aug 6 • 40
Self-Supervised Learning of Structured Dynamics from Videos Paper • 2607.21576 • Published Jul 23 • 21
TimeLens2: Generalist Video Temporal Grounding with Multimodal LLMs Paper • 2607.17423 • Published Jul 19 • 101
Why Can't I Open My Drawer? Mitigating Object-Driven Shortcuts in Zero-Shot Compositional Action Recognition Paper • 2601.16211 • Published Jul 2 • 51
Domain Arithmetic: One-Shot VLA Adaptation under Environmental Shifts Paper • 2607.00666 • Published Jul 1 • 24
Rethinking RAG in Long Videos: What to Retrieve and How to Use It? Paper • 2606.13141 • Published Jun 11 • 35
Video-MME-v2: Towards the Next Stage in Benchmarks for Comprehensive Video Understanding Paper • 2604.05015 • Published Apr 6 • 233
Does Your Reasoning Model Implicitly Know When to Stop Thinking? Paper • 2602.08354 • Published Feb 9 • 139