The patent badge is an abbreviated version of the USPTO patent document. The patent badge does contain a link to the full patent document.

The patent badge is an abbreviated version of the USPTO patent document. The patent badge covers the following: Patent number, Date patent was issued, Date patent was filed, Title of the patent, Applicant, Inventor, Assignee, Attorney firm, Primary examiner, Assistant examiner, CPCs, and Abstract. The patent badge does contain a link to the full patent document (in Adobe Acrobat format, aka pdf). To download or print any patent click here.

Date of Patent:
May. 19, 2026

Filed:

May. 20, 2022
Applicants:

Massachusetts Institute of Technology, Cambridge, MA (US);

Ntt Research, Incorporated, Sunnyvale, CA (US);

Inventors:

Manya Ghobadi, Cambridge, MA (US);

Zhizhen Zhong, Cambridge, MA (US);

Weiyang Wang, Cambridge, MA (US);

Liane Sarah Beland Bernstein, Cambridge, MA (US);

Alexander Sludds, Cambridge, MA (US);

Ryan Hamerly, Cambridge, MA (US);

Dirk Robert Englund, Brookline, MA (US);

Assignees:

Massachusetts Institute of Technology, Cambridge, MA (US);

NTT Research, Incorporated, Sunnyvale, CA (US);

Attorney:
Primary Examiner:
Int. Cl.
CPC ...
H04Q 11/00 (2006.01); G06N 5/04 (2023.01); H04B 10/524 (2013.01);
U.S. Cl.
CPC ...
H04Q 11/0005 (2013.01); G06N 5/04 (2013.01); H04B 10/524 (2013.01); H04Q 2011/0039 (2013.01); H04Q 2011/0041 (2013.01);
Abstract

In-network Optical Inference (IOI) provides low-latency machine learning inference by leveraging programmable switches and optical matrix multiplication. IOI uses a transceiver module, called a Neuro Transceiver, with an optical processor to perform linear operations, such as matrix multiplication, in the optical domain. IOI's transceiver modules can be plugged into programmable packet switches, which are programmed to perform non-linear activations in the electronic domain and to respond to inference queries. Processing inference queries at the programmable packet switches inside the network, without sending them to cloud or edge inference servers, significantly reduces end-to-end inference latency experienced by users.


Find Patent Forward Citations

Loading…