6 October 2022 VP-KLNet: efficient 6D object pose estimation with an enhanced vector-field prediction network and a keypoint localization network
Yaoyin Zhang, Lili Wan, Yazhi Zhu, Wanru Xu, Shenghui Wang
Author Affiliations +
Abstract

6D object pose estimation is essential for many applications with high demands in accuracy and speed. Compared with end-to-end approaches, pixel-wise voting network (PVNet), a vector-field based two-stage approach, has shown the superiority in accuracy but the inferiority in speed because of the time-consuming RANSAC-based voting strategy. To resolve this problem, we propose an efficient deep architecture that consists of an enhanced vector-field prediction network (VPNet) and a keypoint localization network (KLNet) and call it VP-KLNet. Specifically, the KLNet replaces PVNet’s time-consuming voting scheme by directly regressing 2D keypoints from the vector fields, which significantly improves the inference speed. Furthermore, to capture multiscale contextual information, we embed the pyramid pooling module between the encoder and decoder in VPNet to obtain more accurate object segmentation and vector-field prediction, with negligible speed loss. Experiments demonstrate that our method has more than 50% improved in the running speed to the baseline method PVNet and achieves comparable accuracy with the state-of-the-art methods on the LINEMOD and occlusion LINEMOD datasets.

© 2022 SPIE and IS&T
Yaoyin Zhang, Lili Wan, Yazhi Zhu, Wanru Xu, and Shenghui Wang "VP-KLNet: efficient 6D object pose estimation with an enhanced vector-field prediction network and a keypoint localization network," Journal of Electronic Imaging 31(5), 053026 (6 October 2022). https://doi.org/10.1117/1.JEI.31.5.053026
Received: 22 April 2022; Accepted: 20 September 2022; Published: 6 October 2022
Advertisement
Advertisement
RIGHTS & PERMISSIONS
Get copyright permission  Get copyright permission on Copyright Marketplace
KEYWORDS
Visualization

Network architectures

Image segmentation

3D modeling

Convolution

Data modeling

Computer programming

RELATED CONTENT


Back to Top