Beyond Data Scaling: Representation-Centric Continued Pre-training for Vision-Language-Action Models
Paper • 2608.27550 • Published • 89
None defined yet.
Utonia: Toward One Encoder for All Point Clouds
Concerto: Joint 2D-3D Self-Supervised Learning Emerges Spatial Representations