BshapeNet: Object Detection and Instance Segmentation with Bounding   Shape Masks

Ba Rom Kang; Ha Young Kim

arXiv:1810.10327·cs.CV·August 1, 2019

BshapeNet: Object Detection and Instance Segmentation with Bounding Shape Masks

Ba Rom Kang, Ha Young Kim

PDF

TL;DR

This paper introduces BshapeNet+ which enhances object detection and instance segmentation by incorporating boundary shape masks, leading to significant performance improvements over existing models on standard benchmarks.

Contribution

The paper proposes novel boundary shape masks and integrates them into detection and segmentation frameworks, achieving state-of-the-art results.

Findings

01

BshapeNet+ outperforms Faster R-CNN+RoIAlign in detection AP.

02

It achieves 24.9% AP on small objects in COCO.

03

Substantially better instance segmentation than Mask R-CNN.

Abstract

Recent object detectors use four-coordinate bounding box (bbox) regression to predict object locations. Providing additional information indicating the object positions and coordinates will improve detection performance. Thus, we propose two types of masks: a bbox mask and a bounding shape (bshape) mask, to represent the object's bbox and boundary shape, respectively. For each of these types, we consider two variants: the Thick model and the Scored model, both of which have the same morphology but differ in ways to make their boundaries thicker. To evaluate the proposed masks, we design extended frameworks by adding a bshape mask (or a bbox mask) branch to a Faster R-CNN framework, and call this BshapeNet (or BboxNet). Further, we propose BshapeNet+, a network that combines a bshape mask branch with a Mask R-CNN to improve instance segmentation as well as detection. Among our proposed…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

MethodsRegion Proposal Network · Softmax · Convolution · RoIPool · Faster R-CNN · RoIAlign · Mask R-CNN