
CH
Chunming He, Rihan Zhang, Lei Xu, Guanyi Qin, Chengyu Fang, Longxiang Tang, Fengyang Xiao, Sina Farsiu
· 1 min read
ResearcharXiv cs.CV
UnfoldCRF: Structured Mask Refinement with Image-Conditioned Latent Regions
arXiv:2609.33996v1 Announce Type: new
Abstract: Learned mask refiners improve segmentation accuracy, but it is hard to tell how much of the improvement comes from explicit structure rather than from extra capacity, and whether it holds up when the mask generator or its error distribution changes. UnfoldCRF treats refinement as inference in a conditional random field over pixel labels and latent region variables. Its energy has a corrected unary term, learned local pairwise interactions, and image-conditioned latent-region consistency, with a null state that lets a region with weak label agreement withdraw from the consistency term; inference unrolls damped mean-field updates on this one energy. To isolate the effect of structure, we compare against recurrent black-box refiners that read the same inputs and receive the same parameter budget, stage count, and supervision. On COD10K, UnfoldCRF beats the strongest matched control by 1.0 $F^\omega_\beta$ point, improves all four COD metrics, and lowers the fraction of images made worse from 11.7\% to 8.5\%. Under a train-once protocol over five datasets and several mask sources, the 2.6M-parameter variant gains 4.2 mean $\Delta$IoU against 2.0 for its control, and a variant built on frozen DINOv2 features matches the strongest foundation-model refiner with about a seventh of its resident parameters while staying ahead of its own control. On mask generators never seen in training, the gain is 2.0 $F^\omega_\beta$ points against 0.6 for the control. Zeroing individual messages shows where the corrections come from: the pairwise messages mostly fix boundaries, the region messages mostly fix non-boundary errors. Code and supporting materials will be publicly released.
Original source
This story was published by arXiv cs.CV and written by Chunming He, Rihan Zhang, Lei Xu, Guanyi Qin, Chengyu Fang, Longxiang Tang, Fengyang Xiao, Sina Farsiu. SyncAI.news shows a preview; the complete article is on the publisher's site.
Read the full story on arxiv.org


