3D Object Detection and Pose Estimation of Unseen Objects in Color Images with Local Surface Embeddings

Giorgia Pitteri; Aurélie Bugeau; Slobodan Ilic; Vincent Lepetit

3D Object Detection and Pose Estimation of Unseen Objects in Color Images with Local Surface Embeddings

Giorgia Pitteri, Aurélie Bugeau, Slobodan Ilic, Vincent Lepetit

Institute of Computer Graphics and Vision (7100)

Research output: Contribution to conference › Paper › peer-review

Abstract

We present an approach for detecting and estimating the 3D poses of objects in images that requires only an untextured CAD model and no training phase for new objects. Our approach combines Deep Learning and 3D geometry: It relies on an embedding of local 3D geometry to match the CAD models to the input images. For points at the surface of objects, this embedding can be computed directly from the CAD model; for image locations, we learn to predict it from the image itself. This establishes correspondences between 3D points on the CAD model and 2D locations of the input images. However, many of these correspondences are ambiguous as many points may have similar local geometries. We show that we can use Mask-RCNN in a class-agnostic way to detect the new objects without retraining and thus drastically limit the number of possible correspondences. We can then robustly estimate a 3D pose from these discriminative correspondences using a RANSAC-like algorithm. We demonstrate the performance of this approach on the T-LESS dataset, by using a small number of objects to learn the embedding and testing it on the other objects. Our experiments show that our method is on par or better than previous methods.

Original language	English
Publication status	Published - 2020
Event	2020 Asian Conference on Computer Vision: ACCV 2020 - Virtuell, Japan Duration: 30 Nov 2020 → 4 Dec 2020

Conference

Conference	2020 Asian Conference on Computer Vision
Abbreviated title	ACCV 2020
Country/Territory	Japan
City	Virtuell
Period	30/11/20 → 4/12/20

Access to Document

https://openaccess.thecvf.com/content/ACCV2020/papers/Pitteri_3D_Object_Detection_and_Pose_Estimation_of_Unseen_Objects_in_ACCV_2020_paper.pdf

Cite this

Pitteri, G., Bugeau, A., Ilic, S., & Lepetit, V. (2020). 3D Object Detection and Pose Estimation of Unseen Objects in Color Images with Local Surface Embeddings. Paper presented at 2020 Asian Conference on Computer Vision, Virtuell, Japan. https://openaccess.thecvf.com/content/ACCV2020/papers/Pitteri_3D_Object_Detection_and_Pose_Estimation_of_Unseen_Objects_in_ACCV_2020_paper.pdf

Pitteri, G, Bugeau, A, Ilic, S & Lepetit, V 2020, '3D Object Detection and Pose Estimation of Unseen Objects in Color Images with Local Surface Embeddings', Paper presented at 2020 Asian Conference on Computer Vision, Virtuell, Japan, 30/11/20 - 4/12/20. <https://openaccess.thecvf.com/content/ACCV2020/papers/Pitteri_3D_Object_Detection_and_Pose_Estimation_of_Unseen_Objects_in_ACCV_2020_paper.pdf>

@conference{49b587357ab543159674e01f502522ce,

title = "3D Object Detection and Pose Estimation of Unseen Objects in Color Images with Local Surface Embeddings",

abstract = "We present an approach for detecting and estimating the 3D poses of objects in images that requires only an untextured CAD model and no training phase for new objects. Our approach combines Deep Learning and 3D geometry: It relies on an embedding of local 3D geometry to match the CAD models to the input images. For points at the surface of objects, this embedding can be computed directly from the CAD model; for image locations, we learn to predict it from the image itself. This establishes correspondences between 3D points on the CAD model and 2D locations of the input images. However, many of these correspondences are ambiguous as many points may have similar local geometries. We show that we can use Mask-RCNN in a class-agnostic way to detect the new objects without retraining and thus drastically limit the number of possible correspondences. We can then robustly estimate a 3D pose from these discriminative correspondences using a RANSAC-like algorithm. We demonstrate the performance of this approach on the T-LESS dataset, by using a small number of objects to learn the embedding and testing it on the other objects. Our experiments show that our method is on par or better than previous methods.",

author = "Giorgia Pitteri and Aur{\'e}lie Bugeau and Slobodan Ilic and Vincent Lepetit",

year = "2020",

language = "English",

note = "2020 Asian Conference on Computer Vision : ACCV 2020, ACCV 2020 ; Conference date: 30-11-2020 Through 04-12-2020",

}

TY - CONF

T1 - 3D Object Detection and Pose Estimation of Unseen Objects in Color Images with Local Surface Embeddings

AU - Pitteri, Giorgia

AU - Bugeau, Aurélie

AU - Ilic, Slobodan

AU - Lepetit, Vincent

PY - 2020

Y1 - 2020

N2 - We present an approach for detecting and estimating the 3D poses of objects in images that requires only an untextured CAD model and no training phase for new objects. Our approach combines Deep Learning and 3D geometry: It relies on an embedding of local 3D geometry to match the CAD models to the input images. For points at the surface of objects, this embedding can be computed directly from the CAD model; for image locations, we learn to predict it from the image itself. This establishes correspondences between 3D points on the CAD model and 2D locations of the input images. However, many of these correspondences are ambiguous as many points may have similar local geometries. We show that we can use Mask-RCNN in a class-agnostic way to detect the new objects without retraining and thus drastically limit the number of possible correspondences. We can then robustly estimate a 3D pose from these discriminative correspondences using a RANSAC-like algorithm. We demonstrate the performance of this approach on the T-LESS dataset, by using a small number of objects to learn the embedding and testing it on the other objects. Our experiments show that our method is on par or better than previous methods.

AB - We present an approach for detecting and estimating the 3D poses of objects in images that requires only an untextured CAD model and no training phase for new objects. Our approach combines Deep Learning and 3D geometry: It relies on an embedding of local 3D geometry to match the CAD models to the input images. For points at the surface of objects, this embedding can be computed directly from the CAD model; for image locations, we learn to predict it from the image itself. This establishes correspondences between 3D points on the CAD model and 2D locations of the input images. However, many of these correspondences are ambiguous as many points may have similar local geometries. We show that we can use Mask-RCNN in a class-agnostic way to detect the new objects without retraining and thus drastically limit the number of possible correspondences. We can then robustly estimate a 3D pose from these discriminative correspondences using a RANSAC-like algorithm. We demonstrate the performance of this approach on the T-LESS dataset, by using a small number of objects to learn the embedding and testing it on the other objects. Our experiments show that our method is on par or better than previous methods.

M3 - Paper

T2 - 2020 Asian Conference on Computer Vision

Y2 - 30 November 2020 through 4 December 2020

ER -

3D Object Detection and Pose Estimation of Unseen Objects in Color Images with Local Surface Embeddings

Abstract

Conference

Access to Document

Fingerprint

Cite this