Let's Scale Step by Step: Compute-Efficient Hyperparameter Transfer for Large-Scale Mixture-of-Experts Paper • 2608.20061 • Published 9 days ago • 44
view article Article What We Learned by Reproducing 2,200 papers from ICML abidlabs • 17 days ago • 105
BDH-CQ: In-Context Learning with Recurrent Latent Reasoning Paper • 2608.09888 • Published 20 days ago • 762
nyralabs/CrisperWhisper2.0_large Automatic Speech Recognition • 2B • Updated 16 days ago • 30.3k • 107
Team RAS in 11th ABAW Competition: Multimodal Ambivalence Recognition Approach Paper • 2607.14702 • Published Jul 16
Team LEYA in 10th ABAW Competition: Multimodal Ambivalence/Hesitancy Recognition Approach Paper • 2603.12848 • Published Mar 13
VideoChat3: Fully Open Video MLLM for Efficient and Generalist Video Understanding Paper • 2607.14935 • Published Jul 16 • 172