The patent badge is an abbreviated version of the USPTO patent document. The patent badge does contain a link to the full patent document.

The patent badge is an abbreviated version of the USPTO patent document. The patent badge covers the following: Patent number, Date patent was issued, Date patent was filed, Title of the patent, Applicant, Inventor, Assignee, Attorney firm, Primary examiner, Assistant examiner, CPCs, and Abstract. The patent badge does contain a link to the full patent document (in Adobe Acrobat format, aka pdf). To download or print any patent click here.

Date of Patent:
Aug. 13, 2024

Filed:

Jul. 26, 2022
Applicant:

Microsoft Technology Licensing, Llc, Redmond, WA (US);

Inventors:

Wei Liu, Beijing, CN;

Padma Varadharajan, San Jose, CA (US);

Piyush Behre, Santa Clara, CA (US);

Nicholas Kibre, Redwood City, CA (US);

Edward C. Lin, Beijing, CN;

Shuangyu Chang, Davis, CA (US);

Che Zhao, Beijing, CN;

Khuram Shahid, Woodinville, WA (US);

Heiko Willy Rahmel, Bellevue, WA (US);

Assignee:
Attorney:
Primary Examiner:
Int. Cl.
CPC ...
G06F 40/284 (2020.01); G06F 40/117 (2020.01); G06F 40/151 (2020.01); G06F 40/166 (2020.01);
U.S. Cl.
CPC ...
G06F 40/151 (2020.01); G06F 40/117 (2020.01); G06F 40/166 (2020.01); G06F 40/284 (2020.01);
Abstract

Solutions for custom display post processing (DPP) in speech recognition (SR) use a customized multi-stage DPP pipeline that transforms a stream of SR tokens from lexical form to display form. A first transformation stage of the DPP pipeline receives the stream of tokens, in turn, by an upstream filter, a base model stage, and a downstream filter, and transforms a first aspect of the stream of tokens (e.g., disfluency, inverse text normalization (ITN), capitalization, etc.) from lexical form into display form. The upstream filter and/or the downstream filter alter the stream of tokens to change the default behavior of the DPP pipeline into custom behavior. Additional transformation stages of the DPP pipeline perform further transforms, allowing for outputting final text in a display format that is customized for a specific user. This permits each user to efficiently leverage a common baseline DPP pipeline to produce a custom output.


Find Patent Forward Citations

Loading…