04:00
2026-09-07
arxiv.org
computer-vision
DART: Depth-as-Target Pretraining for Surgical Vision Foundation Models
Researchers introduced DART, an RGB-D pretraining method that builds on DINOv2 by adding a pixel-space depth reconstruction objective supervised by pseudo-labeled depth, improving vision foundation moβ¦