Detecting 11K Classes: Large Scale Object Detection without Fine-Grained Bounding Boxes

Yang, Hao; Wu, Hao; Chen, Hao

Computer Science > Computer Vision and Pattern Recognition

arXiv:1908.05217 (cs)

[Submitted on 14 Aug 2019]

Title:Detecting 11K Classes: Large Scale Object Detection without Fine-Grained Bounding Boxes

Authors:Hao Yang, Hao Wu, Hao Chen

View PDF

Abstract:Recent advances in deep learning greatly boost the performance of object detection. State-of-the-art methods such as Faster-RCNN, FPN and R-FCN have achieved high accuracy in challenging benchmark datasets. However, these methods require fully annotated object bounding boxes for training, which are incredibly hard to scale up due to the high annotation cost. Weakly-supervised methods, on the other hand, only require image-level labels for training, but the performance is far below their fully-supervised counterparts. In this paper, we propose a semi-supervised large scale fine-grained detection method, which only needs bounding box annotations of a smaller number of coarse-grained classes and image-level labels of large scale fine-grained classes, and can detect all classes at nearly fully-supervised accuracy. We achieve this by utilizing the correlations between coarse-grained and fine-grained classes with shared backbone, soft-attention based proposal re-ranking, and a dual-level memory module. Experiment results show that our methods can achieve close accuracy on object detection to state-of-the-art fully-supervised methods on two large scale datasets, ImageNet and OpenImages, with only a small fraction of fully annotated classes.

Comments:	Accepted to ICCV 2019
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:1908.05217 [cs.CV]
	(or arXiv:1908.05217v1 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.1908.05217

Submission history

From: Hao Yang Dr [view email]
[v1] Wed, 14 Aug 2019 16:42:15 UTC (3,402 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Detecting 11K Classes: Large Scale Object Detection without Fine-Grained Bounding Boxes

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Detecting 11K Classes: Large Scale Object Detection without Fine-Grained Bounding Boxes

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators