Let's Scale Step by Step: Compute-Efficient Hyperparameter Transfer for Large-Scale Mixture-of-Experts Paper • 2608.20061 • Published 18 days ago • 46
view article Article What We Learned by Reproducing 2,200 papers from ICML abidlabs • 26 days ago • 110
BDH-CQ: In-Context Learning with Recurrent Latent Reasoning Paper • 2608.09888 • Published 29 days ago • 776
nyralabs/CrisperWhisper2.0_large Automatic Speech Recognition • 2B • Updated 25 days ago • 20.2k • 111
Team RAS in 11th ABAW Competition: Multimodal Ambivalence Recognition Approach Paper • 2607.14702 • Published Jul 16
Team LEYA in 10th ABAW Competition: Multimodal Ambivalence/Hesitancy Recognition Approach Paper • 2603.12848 • Published Mar 13
VideoChat3: Fully Open Video MLLM for Efficient and Generalist Video Understanding Paper • 2607.14935 • Published Jul 16 • 172