The patent badge is an abbreviated version of the USPTO patent document. The patent badge does contain a link to the full patent document.

The patent badge is an abbreviated version of the USPTO patent document. The patent badge covers the following: Patent number, Date patent was issued, Date patent was filed, Title of the patent, Applicant, Inventor, Assignee, Attorney firm, Primary examiner, Assistant examiner, CPCs, and Abstract. The patent badge does contain a link to the full patent document (in Adobe Acrobat format, aka pdf). To download or print any patent click here.

Date of Patent:
Mar. 10, 2026

Filed:

Oct. 24, 2023
Applicant:

Google Llc, Mountain View, CA (US);

Inventors:

Nobuyuki Morioka, Mountain View, CA (US);

Byungha Chun, Tokyo, JP;

Nanxin Chen, Mountain View, CA (US);

Yu Zhang, Mountain View, CA (US);

Yifan Ding, Mountain View, CA (US);

Assignee:

Google LLC, Mountain View, CA (US);

Attorneys:
Primary Examiner:
Int. Cl.
CPC ...
G10L 13/027 (2013.01); G06N 3/04 (2023.01); G06N 3/045 (2023.01); G06N 3/0455 (2023.01); G06N 20/10 (2019.01); G10L 13/02 (2013.01); G10L 13/033 (2013.01); G10L 13/047 (2013.01); G10L 13/08 (2013.01);
U.S. Cl.
CPC ...
G10L 13/027 (2013.01); G06N 3/04 (2013.01); G06N 3/045 (2023.01); G06N 3/0455 (2023.01); G06N 20/10 (2019.01); G10L 13/02 (2013.01); G10L 13/033 (2013.01); G10L 13/047 (2013.01); G10L 13/08 (2013.01);
Abstract

A method for residual adapters for few-shot text-to-speech speaker adaptation includes obtaining a text-to-speech (TTS) model configured to convert text into representations of synthetic speech, the TTS model pre-trained on an initial training data set. The method further includes augmenting the TTS model with a stack of residual adapters. The method includes receiving an adaption training data set including one or more spoken utterances spoken by a target speaker, each spoken utterance in the adaptation training data set paired with corresponding input text associated with a transcription of the spoken utterance. The method also includes adapting, using the adaption training data set, the TTS model augmented with the stack of residual adapters to learn how to synthesize speech in a voice of the target speaker by optimizing the stack of residual adapters while parameters of the TTS model are frozen.


Find Patent Forward Citations

Loading…