The patent badge is an abbreviated version of the USPTO patent document. The patent badge does contain a link to the full patent document.

The patent badge is an abbreviated version of the USPTO patent document. The patent badge covers the following: Patent number, Date patent was issued, Date patent was filed, Title of the patent, Applicant, Inventor, Assignee, Attorney firm, Primary examiner, Assistant examiner, CPCs, and Abstract. The patent badge does contain a link to the full patent document (in Adobe Acrobat format, aka pdf). To download or print any patent click here.

Date of Patent:
Apr. 28, 2026

Filed:

Jul. 07, 2022
Applicant:

Nvidia Corporation, Santa Clara, CA (US);

Inventors:

Yeongho Seol, Seoul, KR;

Simon Yuen, Playa Vista, CA (US);

Dmitry Aleksandrovich Korobchenko, Moscow, RU;

Mingquan Zhou, Millbrae, CA (US);

Ronan Browne, Fairfax, CA (US);

Wonmin Byeon, Santa Cruz, CA (US);

Assignee:

Nvidia Corporation, Santa Clara, CA (US);

Attorney:
Primary Examiner:
Assistant Examiner:
Int. Cl.
CPC ...
G06T 13/20 (2011.01); G06T 13/40 (2011.01); G06T 17/20 (2006.01); G10L 15/16 (2006.01); G10L 21/10 (2013.01); G10L 25/63 (2013.01);
U.S. Cl.
CPC ...
G06T 13/205 (2013.01); G06T 13/40 (2013.01); G06T 17/20 (2013.01); G10L 15/16 (2013.01); G10L 21/10 (2013.01); G10L 25/63 (2013.01);
Abstract

A deep neural network can be trained to output motion or deformation information for a character that is representative of the character uttering speech contained in audio input, which is accurate for an emotional state of the character. The character can have different facial components or regions (e.g., head, skin, eyes, tongue) modeled separately, such that the network can output motion or deformation information for each of these different facial components. During training, the network can be provided with emotion and/or style vectors that indicate information to be used in generating realistic animation for input speech, as may relate to one or more emotions to be exhibited by the character, a relative weighting of those emotions, and any style or adjustments to be made to how the character expresses that emotional state. The network output can be provided to a renderer to generate audio-driven facial animation that is emotion-accurate.


Find Patent Forward Citations

Loading…