When EOS Tokens Disagree: Understanding Length Inflation in On-Policy Distillation Paper • 2609.20511 • Published 1 day ago • 10
ModularRSI: Modular and Generalizable Recursive Harness Self-Improvement Paper • 2609.14857 • Published 5 days ago • 87
Occamy-1.0: Open Pareto-frontier 35B Intelligence for Co-work Paper • 2609.11977 • Published 15 days ago • 110
Running Featured 789 Agent Memory Leaderboard 🧠789 Unified memory evaluation · Results expected August 12.
ZGCM-1: A Fully Open and Extremely Efficient Foundation Model for Math and Agentic Search Paper • 2609.13356 • Published 8 days ago • 249
Dream-RSI: Recursive Self-Improvement through Evolving Worlds Paper • 2609.14858 • Published 5 days ago • 222
DataFlex-RL: An Evaluation Platform for RLVR Data Policies Paper • 2609.06107 • Published 14 days ago • 80
meta-llama/Llama-3.1-8B-Instruct Text Generation • 8B • Updated Sep 25, 2024 • 5.93M • • 7.71k
SceneMosaic: Efficient and Diverse Simulation-Ready Scene Generation via Hybrid Agentic Layout Evolution Paper • 2609.05594 • Published 15 days ago • 36