The patent badge is an abbreviated version of the USPTO patent document. The patent badge does contain a link to the full patent document.

The patent badge is an abbreviated version of the USPTO patent document. The patent badge covers the following: Patent number, Date patent was issued, Date patent was filed, Title of the patent, Applicant, Inventor, Assignee, Attorney firm, Primary examiner, Assistant examiner, CPCs, and Abstract. The patent badge does contain a link to the full patent document (in Adobe Acrobat format, aka pdf). To download or print any patent click here.

Date of Patent:
Jun. 09, 2026

Filed:

Oct. 26, 2023
Applicant:

Google Llc, Mountain View, CA (US);

Inventors:

Weiran Wang, San Jose, CA (US);

Ding Zhao, Mountain View, CA (US);

Shaojin Ding, Mountain View, CA (US);

Hao Zhang, Jericho, NY (US);

Shuo-Yiin Chang, Sunnyvale, CA (US);

David Johannes Rybach, Munich, DE;

Tara N. Sainath, Jersey City, NJ (US);

Yanzhang He, Mountain View, CA (US);

Ian Mcgraw, Mountain View, CA (US);

Shankar Kumar, New York, NY (US);

Assignee:

Google LLC, Mountain View, CA (US);

Attorneys:
Primary Examiner:
Int. Cl.
CPC ...
G10L 15/26 (2006.01); G06F 40/284 (2020.01); G10L 15/06 (2013.01);
U.S. Cl.
CPC ...
G10L 15/063 (2013.01); G06F 40/284 (2020.01); G10L 15/26 (2013.01);
Abstract

A method includes receiving a training dataset that includes one or more spoken training utterances for training an automatic speech recognition (ASR) model. Each spoken training utterance in the training dataset paired with a corresponding transcription and a corresponding target sequence of auxiliary tokens. For each spoken training utterance, the method includes generating a speech recognition hypothesis for a corresponding spoken training utterance, determining a speech recognition loss based on the speech recognition hypothesis and the corresponding transcription, generating a predicted auxiliary token for the corresponding spoken training utterance, and determining an auxiliary task loss based on the predicted auxiliary token and the corresponding target sequence of auxiliary tokens. The method also includes the ASR model jointly on the speech recognition loss and the auxiliary task loss determined for each spoken training utterance.


Find Patent Forward Citations

Loading…