HyperAIHyperAI

Command Palette

Search for a command to run...

4 months ago

PVNet: Pixel-wise Voting Network for 6DoF Pose Estimation

Sida Peng; Yuan Liu; Qixing Huang; Hujun Bao; Xiaowei Zhou

PVNet: Pixel-wise Voting Network for 6DoF Pose Estimation

Abstract

This paper addresses the challenge of 6DoF pose estimation from a single RGB image under severe occlusion or truncation. Many recent works have shown that a two-stage approach, which first detects keypoints and then solves a Perspective-n-Point (PnP) problem for pose estimation, achieves remarkable performance. However, most of these methods only localize a set of sparse keypoints by regressing their image coordinates or heatmaps, which are sensitive to occlusion and truncation. Instead, we introduce a Pixel-wise Voting Network (PVNet) to regress pixel-wise unit vectors pointing to the keypoints and use these vectors to vote for keypoint locations using RANSAC. This creates a flexible representation for localizing occluded or truncated keypoints. Another important feature of this representation is that it provides uncertainties of keypoint locations that can be further leveraged by the PnP solver. Experiments show that the proposed approach outperforms the state of the art on the LINEMOD, Occlusion LINEMOD and YCB-Video datasets by a large margin, while being efficient for real-time pose estimation. We further create a Truncation LINEMOD dataset to validate the robustness of our approach against truncation. The code will be avaliable at https://zju-3dv.github.io/pvnet/.

Code Repositories

zju3dv/pvnet
pytorch
Mentioned in GitHub
ivano-donadi/kvn
pytorch
Mentioned in GitHub
edavalosanaya/FastPoseCNN
pytorch
Mentioned in GitHub

Benchmarks

BenchmarkMethodologyMetrics
6d-pose-estimation-on-linemodPVNet
Accuracy: 99%
Accuracy (ADD): 86.27%
Mean ADD: 86.27
6d-pose-estimation-on-ycb-videoPVNet
Mean AUC: 73.4%
6d-pose-estimation-using-rgb-on-occlusionPVNet
Mean ADD: 40.77

Build AI with AI

From idea to launch — accelerate your AI development with free AI co-coding, out-of-the-box environment and best price of GPUs.

AI Co-coding
Ready-to-use GPUs
Best Pricing
Get Started

Hyper Newsletters

Subscribe to our latest updates
We will deliver the latest updates of the week to your inbox at nine o'clock every Monday morning
Powered by MailChimp