The Embedder's Dilemma: LLMs Are Better, but at What Cost? Paper • 2608.12875 • Published 15 days ago • 14
Thinking in a Low-Resource Language: What SFT Builds, What RL Fixes, What Accuracy Cannot See Paper • 2608.17744 • Published 9 days ago • 15
FM-Bench: A Benchmark for Long-Horizon Management with Competing Agents Paper • 2608.18423 • Published 9 days ago • 20
SimpleOPD: Simple Tokenizer-Agnostic On-Policy Distillation for Long-Context Reasoning Paper • 2608.14277 • Published 14 days ago • 35
view article Article Meet North Micro Vision: A 2.4B Native-Resolution Vision-Language Model CohereLabs • 15 days ago • 38
Massive Activations in Hybrid Linear Attention Large Language Models: Pre-Attention Spikes and Inter-Spike Plateaus Paper • 2608.12149 • Published 16 days ago • 30
LLMRouter: Unified Infrastructure for Developing, Evaluating, and Deploying LLM Routers Paper • 2608.06867 • Published 21 days ago • 109
Poor Man's Agentic Modeling: Simulating Large LLM-Agent Societies on a Laptop Paper • 2608.11215 • Published Jul 19 • 6
view article Article FineBooks: are open OCR models good enough to unlock historical knowledge? finebooks • 18 days ago • 25
UEmbed: Unified Sparse and Dense Multimodal Embeddings Paper • 2608.02583 • Published 25 days ago • 50
Embedding Models Collection Run or fine-tune embedding models with Unsloth. • 14 items • Updated 1 day ago • 9
Safeguards Based on Copyable Context Cannot Provide Reliable Safety for LLMs Paper • 2607.27951 • Published 29 days ago • 7
view article Article Anatomy of a Frontier Lab Agent Intrusion: A Technical Timeline of the July 2026 Incident +2 hlarcher, XciD, raphael-gl, chris-rannou • Jul 27 • 482
Apertus v1.5 Collection 8B and 70B open-data open-weights models, globally multilingual with multimodal input and reasoning support • 2 items • Updated 6 days ago • 31