HERO Collection [CoRL 2026] HERO: Learning Humanoid End-Effector Control for Visual Whole-Body Open-Vocabulary Object Grasping • 3 items • Updated 4 days ago • 1
HERO Collection [CoRL 2026] HERO: Learning Humanoid End-Effector Control for Visual Whole-Body Open-Vocabulary Object Grasping • 3 items • Updated 4 days ago • 1
HERO Collection [CoRL 2026] HERO: Learning Humanoid End-Effector Control for Visual Whole-Body Open-Vocabulary Object Grasping • 3 items • Updated 4 days ago • 1
Humanoid-GPT: Scaling Data and Structure for Zero-Shot Motion Tracking Paper • 2606.03985 • Published Jun 2 • 42
ImageWAM: Do World Action Models Really Need Video Generation, or Just Image Editing? Paper • 2606.19531 • Published Jun 17 • 28
PerceptionRubrics: Calibrating Multimodal Evaluation to Human Perception Paper • 2606.28322 • Published Jun 26 • 42
ULTRA: Unified Multimodal Control for Autonomous Humanoid Whole-Body Loco-Manipulation Paper • 2603.03279 • Published Mar 3 • 1
ULTRA: Unified Multimodal Control for Autonomous Humanoid Whole-Body Loco-Manipulation Paper • 2603.03279 • Published Mar 3 • 1
Learning Humanoid End-Effector Control for Open-Vocabulary Visual Loco-Manipulation Paper • 2602.16705 • Published Feb 18 • 26
Learning Humanoid End-Effector Control for Open-Vocabulary Visual Loco-Manipulation Paper • 2602.16705 • Published Feb 18 • 26
Learning Humanoid End-Effector Control for Open-Vocabulary Visual Loco-Manipulation Paper • 2602.16705 • Published Feb 18 • 26
DINOv3 Collection DINOv3: foundation models producing excellent dense features, outperforming SotA w/o fine-tuning - https://arxiv.org/abs/2508.10104 • 15 items • Updated Mar 10 • 808