Is Next-Chunk Reasoning RL Really Better than SFT? Revisiting Training Strategies under no-CoT Data Paper β’ 2608.23256 β’ Published 27 days ago β’ 22
SciExplore: Evaluating Autonomous Agents from Scientific Navigation to Information Integration Paper β’ 2607.20926 β’ Published Jul 23 β’ 2
AdvancedMathBench: A Benchmark Suite for Advanced Mathematical Proof Generation and Verification Paper β’ 2607.11849 β’ Published Jul 13 β’ 29
Scalable Visual Pretraining for Language Intelligence Paper β’ 2607.09657 β’ Published Jul 10 β’ 56
ThoughtFold: Folding Reasoning Chains via Introspective Preference Learning Paper β’ 2606.03503 β’ Published Jun 2 β’ 25
Intern-S1-Pro: Scientific Multimodal Foundation Model at Trillion Scale Paper β’ 2603.25040 β’ Published Mar 26 β’ 132
Long-horizon Reasoning Agent for Olympiad-Level Mathematical Problem Solving Paper β’ 2512.10739 β’ Published Dec 11, 2025 β’ 47
Achieving Olympia-Level Geometry Large Language Model Agent via Complexity Boosting Reinforcement Learning Paper β’ 2512.10534 β’ Published Dec 11, 2025 β’ 33
OPV: Outcome-based Process Verifier for Efficient Long Chain-of-Thought Verification Paper β’ 2512.10756 β’ Published Dec 11, 2025 β’ 36
ARM-Thinker: Reinforcing Multimodal Generative Reward Models with Agentic Tool Use and Visual Reasoning Paper β’ 2512.05111 β’ Published Dec 4, 2025 β’ 50
Intern-S1: A Scientific Multimodal Foundation Model Paper β’ 2508.15763 β’ Published Aug 21, 2025 β’ 275
MIG: Automatic Data Selection for Instruction Tuning by Maximizing Information Gain in Semantic Space Paper β’ 2504.13835 β’ Published Apr 18, 2025 β’ 39
InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models Paper β’ 2504.10479 β’ Published Apr 14, 2025 β’ 312
Creation-MMBench: Assessing Context-Aware Creative Intelligence in MLLM Paper β’ 2503.14478 β’ Published Mar 18, 2025 β’ 48
Running on Zero Agents 104 Make It Animatable π 104 Authoring Animation-Ready 3D Characters with One Click
Running Agents Featured 136 Open VLM Video Leaderboard π 136 VLMEvalKit Eval Results in video understanding benchmark
HumanVid: Demystifying Training Data for Camera-controllable Human Image Animation Paper β’ 2407.17438 β’ Published Jul 24, 2024 β’ 26