The patent badge is an abbreviated version of the USPTO patent document. The patent badge does contain a link to the full patent document.

The patent badge is an abbreviated version of the USPTO patent document. The patent badge covers the following: Patent number, Date patent was issued, Date patent was filed, Title of the patent, Applicant, Inventor, Assignee, Attorney firm, Primary examiner, Assistant examiner, CPCs, and Abstract. The patent badge does contain a link to the full patent document (in Adobe Acrobat format, aka pdf). To download or print any patent click here.

Date of Patent:
May. 13, 2025

Filed:

Nov. 09, 2022
Applicant:

Beijing Baidu Netcom Science Technology Co., Ltd., Beijing, CN;

Inventors:

Shangwen Lyu, Beijing, CN;

Hongyu Li, Beijing, CN;

Jing Liu, Beijing, CN;

Hua Wu, Beijing, CN;

Haifeng Wang, Beijing, CN;

Attorney:
Primary Examiner:
Int. Cl.
CPC ...
G06V 30/19 (2022.01); G06F 40/205 (2020.01); G06V 30/194 (2022.01); G06V 30/412 (2022.01);
U.S. Cl.
CPC ...
G06V 30/1918 (2022.01); G06F 40/205 (2020.01); G06V 30/194 (2022.01); G06V 30/412 (2022.01);
Abstract

A method for training a document reading comprehension model includes: acquiring a question sample and a rich-text document sample, in which the rich-text document sample includes a real answer of the question sample; acquiring text information and layout information of the rich-text document sample by performing OCR processing on image information of the rich-text document sample; acquiring a predicted answer of the question sample by inputting the text information, the layout information and the image information of the rich-text document sample into a preset reading comprehension model; and training the reading comprehension model based on the real answer and the predicted answer. The method may enhance comprehension ability of the reading comprehension model to the long rich-text document, and save labor cost.


Find Patent Forward Citations

Loading…