The patent badge is an abbreviated version of the USPTO patent document. The patent badge does contain a link to the full patent document.

The patent badge is an abbreviated version of the USPTO patent document. The patent badge covers the following: Patent number, Date patent was issued, Date patent was filed, Title of the patent, Applicant, Inventor, Assignee, Attorney firm, Primary examiner, Assistant examiner, CPCs, and Abstract. The patent badge does contain a link to the full patent document (in Adobe Acrobat format, aka pdf). To download or print any patent click here.

Date of Patent:
Jul. 21, 2026

Filed:

Apr. 25, 2024
Applicants:

Naver Corporation, Seongnam-si, KR;

Naver Labs Corporation, Seongnam-si, KR;

Inventors:

Jérome Revaud, Meylan, FR;

Romain Brégier, Meylan, FR;

Yohann Cabon, Meylan, FR;

Philippe Weinzaepfel, Meylan, FR;

Jongmin Lee, Meylan, FR;

Assignee:
Attorney:
Primary Examiner:
Int. Cl.
CPC ...
G06T 7/73 (2017.01); G06T 7/50 (2017.01); G06V 10/774 (2022.01); G06V 10/80 (2022.01); G06V 10/82 (2022.01); G06V 20/70 (2022.01);
U.S. Cl.
CPC ...
G06T 7/74 (2017.01); G06T 7/50 (2017.01); G06V 10/774 (2022.01); G06V 10/806 (2022.01); G06V 10/82 (2022.01); G06V 20/70 (2022.01); G06T 2207/20081 (2013.01); G06T 2207/20084 (2013.01); G06T 2207/30244 (2013.01);
Abstract

A computer implemented method and system using an object-agnostic model for predicting a pose of an object in an image receives a query image having a target object therein; receives a set of reference images of the target object from different viewpoints; encodes, using a vision transformer, the received query image and the received set of reference images to generate a set of token features for the received query image and a set of token features for the received set of reference images; extracts, using a transformer decoder, information from the set of token features for the encoded reference images with respect to a set of token features for the received query image; processes, using a prediction head, the combined set of token features to generate a 2D-3D mapping and a confidence map of the query image; and processes the 2D-3D mapping and confidence map to determine the pose of the target object in the query image.


Find Patent Forward Citations

Loading…