Return
Deformable Template Network (DTN) for Object Detection
DOI:10.1109/TMM.2021.3075323.png)
Abstract
En 中文
Objects often have different appearances because of viewpoint changes or part deformation. How to reasonably model these variations is still a big challenge for object detection. In this paper, we propose a novel Deformable Template Network (DTN), which exploits the pictorial structure to model possible variations of an object. DTN represents an object by virtue of a generated template in a deformable way. It has two key modules: the template generating module and the part matching module. The template generating module produces a template for a given object which defines the anchor positions of the k X k parts. Based on such a template, the part matching module aims to perform part alignment around the anchor positions. In terms of each part, the matching process makes a trade-off between maximizing the detection score and minimizing the deformation cost relative to the anchor position. Moreover, DTN is a fully convolutional network which means it is competitive in terms of detection efficiency. We evaluate DTN on both the PASCAL VOC and MSCOCO datasets, achieving the state-of-the-art results, an accuracy of 82.7% for PASCAL VOC and of 44.9% for MSCOCO.
Keywords:
object detection
deformable template
part matching
deformation cost
AI Summary
Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.
Journal
IF:
9.7
Papers:
4.5K
Citations:
2.4W
Organization
Cited Papers
Conductivity Enhancement in Thin Silicon-on-Insulator Layer Embedding Artificial Dislocation Network
Thermal comfort, perceived air quality, and cognitive performance when personally controlled air movement is used by tropically acclimatized persons
Indoor Air
IF0

