The patent badge is an abbreviated version of the USPTO patent document. The patent badge does contain a link to the full patent document.

The patent badge is an abbreviated version of the USPTO patent document. The patent badge covers the following: Patent number, Date patent was issued, Date patent was filed, Title of the patent, Applicant, Inventor, Assignee, Attorney firm, Primary examiner, Assistant examiner, CPCs, and Abstract. The patent badge does contain a link to the full patent document (in Adobe Acrobat format, aka pdf). To download or print any patent click here.

Date of Patent:
Feb. 24, 2026

Filed:

Nov. 30, 2021
Applicant:

Samsung Electronics Co., Ltd., Suwon-si, KR;

Inventors:

Myungjong Kim, Milpitas, CA (US);

Vijendra Raj Apsingekar, San Jose, CA (US);

Divya Neelagiri, Dublin, CA (US);

Taeyeon Ki, Milpitas, CA (US);

Assignee:
Attorney:
Primary Examiner:
Int. Cl.
CPC ...
G10L 25/30 (2013.01); G06F 40/263 (2020.01); G06N 3/045 (2023.01); G06N 3/048 (2023.01); G06N 20/00 (2019.01); G10L 15/00 (2013.01); G10L 15/02 (2006.01); G10L 15/06 (2013.01); G10L 15/16 (2006.01); G10L 15/22 (2006.01); G10L 17/02 (2013.01); G10L 17/04 (2013.01); G10L 17/06 (2013.01); G10L 17/18 (2013.01); G10L 21/0272 (2013.01);
U.S. Cl.
CPC ...
G10L 21/0272 (2013.01); G06N 3/045 (2023.01); G06N 3/048 (2023.01); G06N 20/00 (2019.01); G10L 15/005 (2013.01); G10L 15/02 (2013.01); G10L 15/063 (2013.01); G10L 15/22 (2013.01); G10L 17/02 (2013.01); G10L 17/04 (2013.01); G10L 17/06 (2013.01); G10L 17/18 (2013.01); G10L 25/30 (2013.01); G06F 40/263 (2020.01);
Abstract

An apparatus for processing speech data may include a processor configured to: separate speech signals from an input speech; identify a language of each of the speech signals that are separated from the input speech; extract speaker embeddings from the speech signals based on the language of each of the speech signals, using at least one neural network configured to receive the speech signals and output the speaker embeddings; and identify a speaker of each of the speech signals by iteratively clustering the speaker embeddings.


Find Patent Forward Citations

Loading…