NeoHorse-1: Towards Recursive Self-Improvement via Agentic Post-Training with Routing Harness Paper • 2609.08183 • Published 20 days ago • 174
VGI-Bench: Probing Visual Intelligence in Video Generation Models Paper • 2608.19583 • Published Aug 26 • 174
HarnessEval-W: Agentifying the Evaluation of Visual Worlds Paper • 2608.16859 • Published Aug 17 • 122
Fewer Clarifications, Better Code: Benchmarking Cross-Session Personalized Ambiguity Adaptation in Coding Assistants Paper • 2607.26611 • Published Jul 29 • 33
TurboVLA: Real-Time Vision-Language-Action Model at 32 Hz on an RTX 4090 with <1 GB VRAM Paper • 2607.27205 • Published Jul 29 • 69
SANA-Video 2.0: Hybrid Linear Attention with Attention Residuals for Efficient Video Generation Paper • 2607.21553 • Published Jul 23 • 40
Spectral Rewiring for Exploration, Purification, and Model Merging Paper • 2607.03065 • Published Jul 3 • 22
VLA-Corrector: Lightweight Detect-and-Correct Inference for Adaptive Action Horizon Paper • 2607.01804 • Published Jul 2 • 27
Is Position Bias in Dense Retrievers Built In-or Learned from Data? Paper • 2605.26578 • Published May 26 • 17
Geometry Matters: 3D Foundation Priors for Learning Semantic Correspondence Paper • 2605.30093 • Published May 28 • 11
Full Attention Strikes Back: Transferring Full Attention into Sparse within Hundred Training Steps Paper • 2605.16928 • Published May 16 • 91
Agent Bazaar: Enabling Economic Alignment in Multi-Agent Marketplaces Paper • 2605.17698 • Published May 17 • 5
CurveBench: A Benchmark for Exact Topological Reasoning over Nested Jordan Curves Paper • 2605.14068 • Published May 13 • 9
SkillsVote: Lifecycle Governance of Agent Skills from Collection, Recommendation to Evolution Paper • 2605.18401 • Published May 18 • 80
Adam's Law: Textual Frequency Law on Large Language Models Paper • 2604.02176 • Published Apr 2 • 109
GrandCode: Achieving Grandmaster Level in Competitive Programming via Agentic Reinforcement Learning Paper • 2604.02721 • Published Apr 3 • 86
From Blind Spots to Gains: Diagnostic-Driven Iterative Training for Large Multimodal Models Paper • 2602.22859 • Published Feb 26 • 31