The patent badge is an abbreviated version of the USPTO patent document. The patent badge does contain a link to the full patent document.

The patent badge is an abbreviated version of the USPTO patent document. The patent badge covers the following: Patent number, Date patent was issued, Date patent was filed, Title of the patent, Applicant, Inventor, Assignee, Attorney firm, Primary examiner, Assistant examiner, CPCs, and Abstract. The patent badge does contain a link to the full patent document (in Adobe Acrobat format, aka pdf). To download or print any patent click here.

Date of Patent:
Apr. 29, 2025

Filed:

Dec. 20, 2022
Applicant:

Mohamed Bin Zayed University of Artificial Intelligence, Abu Dhabi, AE;

Inventors:

Hanoona Abdul Rasheed Bangalath, Abu Dhabi, AE;

Muhammad Maaz, Abu Dhabi, AE;

Muhammad Uzair Khattak, Abu Dhabi, AE;

Salman Khan, Abu Dhabi, AE;

Fahad Shahbaz Khan, Abu Dhabi, AE;

Attorney:
Primary Examiner:
Int. Cl.
CPC ...
G06V 10/25 (2021.12); G06T 1/00 (2005.12); G06V 10/70 (2021.12); G06V 20/56 (2021.12);
U.S. Cl.
CPC ...
G06V 10/25 (2021.12); G06T 1/0021 (2012.12); G06V 10/70 (2021.12); G06V 20/56 (2021.12); G06V 2201/07 (2021.12);
Abstract

An object detection system and method in which a machine learning engine is configured with a region-based knowledge distillation stage that generates region embeddings from a training image having bounding boxes. A linear layer learns a region-level vision-language mapping for projecting feature embeddings from the training image to a common feature space shared by text embeddings to obtain the region embeddings. An image-level supervision stage generates pseudo-box labels for a classification training image and region embeddings from the training image having bounding boxes and corresponding class labels and the classification training image having an image-level label as input. Pseudo-box labels are determined on the classification training image as an image-level vision-language mapping. A weight transfer function conditions the image-level vision-language mapping on the learned region-level vision-language mapping. A trained object detector outputs a newly captured image annotated with a bounding box for a novel object.


Find Patent Forward Citations

Loading…