Paper recorded by Signals 4 on 2026-09-24 in cs.CV. Abstract reproduced from arXiv; link to the original below.
Published 2026-09-24 on arXiv · recorded by Signals 4 on 2026-09-25
Category: cs.CV · 计算机视觉 · first seen 2026-09-25
Foundation segmenters such as SAM return several plausible masks for an unlabeled image, and a student trained on the wrong one inherits its errors. Choosing among them means querying a second large model or fitting a quality head to annotated masks. We show that a candidate can be judged by what it does to a frozen self-supervised backbone's features. Normalized DINOv2 patch features lie on a hyp