发明名称 Selective display of OCR'ed text and corresponding images from publications on a client device
摘要 Text is extracted from a source image of a publication using an Optical Character Recognition (OCR) process. A document is generated containing text segments of the extracted text. The document includes a control module that responds to user interactions with the displayed document. Responsive to a user selection of a displayed text segment, a corresponding image segment from the source image containing the text is retrieved and rendered in place of the selected text segment. The user can select again to toggle the display back to the text segment. Each text segment can be tagged with a garbage score indicating its quality. If the garbage score of a text segment exceeds a threshold value, the corresponding image segment can be automatically displayed instead.
申请公布号 US9280952(B2) 申请公布日期 2016.03.08
申请号 US201414152893 申请日期 2014.01.10
申请人 Google Inc. 发明人 Ratnakar Viresh;Haugen Frances Bordwell;Popat Ashok
分类号 G09G5/14;G06K9/03;G06T11/00;G06T11/60;G06T11/20 主分类号 G09G5/14
代理机构 Fenwick & West LLP 代理人 Fenwick & West LLP
主权项 1. A computer-implemented method for displaying a document, the method comprising: identifying a document including at least one text segment generated responsive to an Optical Character Recognition (OCR) process performed on an image segment, wherein the text segment includes a plurality of characters; generating a quality score for each of the plurality of characters; generating a segment quality measure for the text segment by averaging the generated quality scores, the segment quality measure indicating a quality of the text segment; and responsive to the segment quality measure not meeting a quality threshold, displaying the image segment instead of the text segment on a display of a client device.
地址 Mountain View CA US