gradients-io-tournaments/tournament-exp-s1-27840c46-dd63-4bfb-abd2-8b4ba556f84a-5Exp3970359e47ff9da0 1B • Updated 1 day ago • 15 • 1
HarnessDev: Can LLMs Create and Evolve Their Own Agent Harness? Paper • 2609.01437 • Published 13 days ago • 267
SafeAtlas-VL: Beyond Binary Multimodal Safety with Large-Scale Data and Guard Models Paper • 2608.29098 • Published 16 days ago • 9
JIT-Agent: Scaling Harness Intelligence via Just-in-Time Harness Evolution Paper • 2608.25593 • Published 19 days ago • 117
VGI-Bench: Probing Visual Intelligence in Video Generation Models Paper • 2608.19583 • Published 19 days ago • 179
Learn What's Left, Not What's Mastered: Saturation Aware Advantage Reweighting for Multi-Reward Policy Optimization Paper • 2608.16072 • Published 28 days ago • 151
Spark-to-Paper: End-to-End Research Paper Generation as a Composable Skill Paper • 2608.11924 • Published Aug 12 • 291
SimWAM: A Simple World Action Model for End-to-End Autonomous Driving Paper • 2608.07468 • Published Aug 7 • 108
Progress Reward Modeling for Robotic Learning: A Comprehensive Survey Paper • 2607.21655 • Published Jul 22 • 193
The Mirage of Optimizing Training Policies: Monotonic Inference Policies as the Real Objective for LLM Reinforcement Learning Paper • 2606.29526 • Published Jun 28 • 171