about
Learning to Segment Every Thing [pdf] (arxiv.org)
2 points by stablemap on Dec 29, 2017 | hide | past | pdf | discuss on HN

In plain words: Instead of drawing masks for every object, it learns from boxes for many categories and masks for a few, using a trick that turns its box-detection knowledge into mask-drawing ability. It segments 3,000 concepts with masks for 80 classes, versus the usual ~100 mask-labeled.

Abstract · Learning to Segment Every Thing

Most methods for object instance segmentation require all training examples to be labeled with segmentation masks. This requirement makes it expensive to annotate new categories and has restricted instance segmentation models to ~100 well-annotated classes. The goal of this paper is to propose a new partially supervised training paradigm, together with a novel weight transfer function, that enables training instance segmentation models on a large set of categories all of which have box annotations, but only a small fraction of which have mask annotations. These contributions allow us to train Mask R-CNN to detect and segment 3000 visual concepts using box annotations from the Visual Genome dataset and mask annotations from the 80 classes in the COCO dataset. We evaluate our approach in a controlled study on the COCO dataset. This work is a first step towards instance segmentation models that have broad comprehension of the visual world.

Ronghang Hu, Piotr Dollár, Kaiming He, Trevor Darrell, Ross Girshick
arXiv:1711.10370 · cs.CV · submitted Nov 28, 2017 · updated Mar 27, 2018
abstract · pdf · html

add comment on HN