Let's Scale Step by Step: Compute-Efficient Hyperparameter Transfer for Large-Scale Mixture-of-Experts Paper • 2608.20061 • Published 10 days ago • 46
view article Article What We Learned by Reproducing 2,200 papers from ICML abidlabs • 18 days ago • 106
BDH-CQ: In-Context Learning with Recurrent Latent Reasoning Paper • 2608.09888 • Published 21 days ago • 764
nyralabs/CrisperWhisper2.0_large Automatic Speech Recognition • 2B • Updated 17 days ago • 30.6k • 108
Team RAS in 11th ABAW Competition: Multimodal Ambivalence Recognition Approach Paper • 2607.14702 • Published Jul 16
Team LEYA in 10th ABAW Competition: Multimodal Ambivalence/Hesitancy Recognition Approach Paper • 2603.12848 • Published Mar 13
VideoChat3: Fully Open Video MLLM for Efficient and Generalist Video Understanding Paper • 2607.14935 • Published Jul 16 • 172