SyncAI.news, a Varaisys broadcasting
Depth-Guided Contrastive Learning for 2D Representations with 3D Spatial Awareness
LZ

Liang Zeng, Maarten Vergauwen

· 1 min read

ResearcharXiv cs.CV

Depth-Guided Contrastive Learning for 2D Representations with 3D Spatial Awareness

arXiv:2609.28159v1 Announce Type: new Abstract: Standard contrastive learning frameworks are mainly designed from a semantic perspective, yet learning 2D visual representations that preserve 3D spatial structure is also important for scene understanding. In this work, we propose Depth-Guided Contrastive Learning (DGCL), a simple auxiliary objective that injects 3D spatial awareness into 2D contrastive representation learning. Our key idea is to use depth to convert local 3D proximity into contrastive similarity: pixels that are closer in 3D space are encouraged to have more similar representations than pixels that are farther apart. Instead of relying on absolute depth values, DGCL formulates supervision through relative 3D distance comparisons among randomly sampled pixels, making the objective invariant to depth scale, efficient to compute, and easy to integrate into existing contrastive frameworks. Experiments across different datasets and models show that DGCL consistently improves 2D representation learning and benefits semantic downstream tasks by stronger spatial and geometric understanding. The code is available on https://github.com/LeungTsang/DGCL.

Original source

This story was published by arXiv cs.CV and written by Liang Zeng, Maarten Vergauwen. SyncAI.news shows a preview; the complete article is on the publisher's site.

Read the full story on arxiv.org

Similar News