The patent badge is an abbreviated version of the USPTO patent document. The patent badge does contain a link to the full patent document.

The patent badge is an abbreviated version of the USPTO patent document. The patent badge covers the following: Patent number, Date patent was issued, Date patent was filed, Title of the patent, Applicant, Inventor, Assignee, Attorney firm, Primary examiner, Assistant examiner, CPCs, and Abstract. The patent badge does contain a link to the full patent document (in Adobe Acrobat format, aka pdf). To download or print any patent click here.

Date of Patent:
Aug. 11, 2026

Filed:

May. 17, 2023
Applicant:

Adobe Inc., San Jose, CA (US);

Inventors:

Bryan Russell, San Francisco, CA (US);

Justin Salamon, San Francisco, CA (US);

Daniel Mckee, Urbana-Champaign, IL (US);

Josef Sivic, Prague, CZ;

Assignee:

Adobe Inc., San Jose, CA (US);

Attorney:
Primary Examiner:
Int. Cl.
CPC ...
G06F 16/432 (2019.01); G06F 16/438 (2019.01); G06F 16/632 (2019.01); G06F 16/68 (2019.01);
U.S. Cl.
CPC ...
G06F 16/4393 (2019.01); G06F 16/434 (2019.01); G06F 16/632 (2019.01); G06F 16/686 (2019.01);
Abstract

Embodiments are disclosed for an audio recommendation system trained to recommend music audio sequences for pairing with query video sequences using neural networks. In particular, in one or more embodiments, the disclosed systems and methods comprise receiving an input including a query video sequence and natural language text. The disclosed systems and methods further comprise generating a fused visual-text embedding based on a visual embedding and a text embedding corresponding to the input. The disclosed systems and methods further comprise comparing audio embeddings for music audio sequences of a music audio sequences database with the fused visual-text embedding. The disclosed systems and methods further comprise determining a music audio sequence from the music audio sequences database as the recommended music audio sequence for pairing with the query video sequence based on a similarity metric calculated between an audio embedding for the music audio sequence and the fused visual-text embedding.


Find Patent Forward Citations

Loading…