Research questionHow can self-supervised surround-view depth estimation remain geometrically consistent when camera views minimally overlap?In surround-camera systems, most pixels lack strong multi-view correspondence and must be interpreted from monocular cues. Different camera intrinsics and limited per-image context can therefore produce inconsistent depth estimates across adjacent views.