Arrow Research search
Back to IJCAI

IJCAI 2017

Saliency Guided End-to-End Learning for Weakly Supervised Object Detection

Conference Paper Machine Learning A-R Artificial Intelligence

Abstract

Weakly supervised object detection (WSOD), which is the problem of learning detectors using only image-level labels, has been attracting more and more interest. However, this problem is quite challenging due to the lack of location supervision. To address this issue, this paper integrates saliency into a deep architecture, in which the location information is explored both explicitly and implicitly. Specifically, we select highly confident object proposals under the guidance of class-specific saliency maps. The location information, together with semantic and saliency information, of the select proposals are then used to explicitly supervise the network by imposing two additional losses. Meanwhile, a saliency prediction sub-network is built in the architecture. The prediction results are used to implicitly guide the localization procedure. The entire network is trained end-to-end. Experiments on PASCAL VOC demonstrate that our approach outperforms all state-of-the-arts.

Authors

Keywords

  • Machine Learning: Deep Learning
  • Machine Learning: Multi-instance/Multi-label/Multi-view learning
  • Robotics and Vision: Vision and Perception

Context

Venue
International Joint Conference on Artificial Intelligence
Archive span
1969-2025
Indexed papers
14525
Paper id
937731255929090249
v2026.09.13