Unsupervised Semantic Segmentation by Distilling Feature Correspondences

2022-03-16ICLR 2022Code Available2· sign in to hype

Mark Hamilton, Zhoutong Zhang, Bharath Hariharan, Noah Snavely, William T. Freeman

Code Available — Be the first to reproduce this paper.

Code

github.com/mhamilton723/STEGO
Officialpytorch★ 785
github.com/leggedrobotics/self_supervised_segmentation
pytorch★ 30
github.com/merantix-momentum/stego-studies
pytorch★ 13

Abstract

Unsupervised semantic segmentation aims to discover and localize semantically meaningful categories within image corpora without any form of annotation. To solve this task, algorithms must produce features for every pixel that are both semantically meaningful and compact enough to form distinct clusters. Unlike previous works which achieve this with a single end-to-end framework, we propose to separate feature learning from cluster compactification. Empirically, we show that current unsupervised feature learning frameworks already generate dense features whose correlations are semantically consistent. This observation motivates us to design STEGO (Self-supervised Transformer with Energy-based Graph Optimization), a novel framework that distills unsupervised features into high-quality discrete semantic labels. At the core of STEGO is a novel contrastive loss function that encourages features to form compact clusters while preserving their relationships across the corpora. STEGO yields a significant improvement over the prior state of the art, on both the CocoStuff (+14 mIoU) and Cityscapes (+9 mIoU) semantic segmentation challenges.

Tasks

Form Semantic Segmentation Unsupervised Semantic Segmentation

Benchmark Results

Dataset	Model	Metric	Claimed	Verified	Status
Cityscapes test	STEGO	mIoU	21	—	Unverified
COCO-Stuff-27	STEGO (ViT-B/8)	Clustering [mIoU]	28.2	—	Unverified
COCO-Stuff-27	STEGO (ViT-S/8)	Clustering [mIoU]	24.5	—	Unverified
Potsdam-3	STEGO	Accuracy	77	—	Unverified

Unsupervised Semantic Segmentation by Distilling Feature Correspondences

Code

Abstract

Tasks

Benchmark Results

Reproductions