Towards Better Object Detection in Scale Variation with Adaptive Feature Selection

December 06, 2020 · Entered Twilight · 🏛 arXiv.org

"Last commit was 5.0 years ago (≥5 year threshold)"

Evidence collected by the PWNC Scanner

Repo contents: .idea, .pre-commit-config.yaml, .style.yapf, .travis.yml, .travis, LICENSE, README.md, configs, contrib, dataset, demo, deploy, docs, ppdet, requirements.txt, slim, tools

Authors Zehui Gong, Dong Li arXiv ID 2012.03265 Category cs.CV: Computer Vision Citations 5 Venue arXiv.org Repository https://github.com/ZeHuiGong/AFSM.git ⭐ 30 Last Checked 2 months ago

Abstract

It is a common practice to exploit pyramidal feature representation to tackle the problem of scale variation in object instances. However, most of them still predict the objects in a certain range of scales based solely or mainly on a single-level representation, yielding inferior detection performance. To this end, we propose a novel adaptive feature selection module (AFSM), to automatically learn the way to fuse multi-level representations in the channel dimension, in a data-driven manner. It significantly improves the performance of the detectors that have a feature pyramid structure, while introducing nearly free inference overhead. Moreover, a class-aware sampling mechanism (CASM) is proposed to tackle the class imbalance problem, by re-weighting the sampling ratio to each of the training images, based on the statistical characteristics of each class. This is crucial to improve the performance of the minor classes. Experimental results demonstrate the effectiveness of the proposed method, with 83.04% mAP at 15.96 FPS on the VOC dataset, and 39.48% AP on the VisDrone-DET validation subset, respectively, outperforming other state-of-the-art detectors considerably. The code is available at https://github.com/ZeHuiGong/AFSM.git.