Generating a combination of a visual query and matching canonical document,申请号US201113309461-传众专利搜索

发明名称	Generating a combination of a visual query and matching canonical document
摘要	A server system receives a visual query from a client system distinct from the server system, performs optical character recognition (OCR) on the visual query to produce text recognition data representing textual characters, including a plurality of textual characters in a contiguous region of the visual query, and scores each textual character in the plurality of textual characters. The server system identifies, in accordance with the scoring, one or more high quality textual strings, each comprising a plurality of high quality textual characters from among the plurality of textual characters in the contiguous region of the visual query; retrieves a canonical document having the one or more high quality textual strings; generates a combination of the visual query and at least a portion of the canonical document; and sends the combination to the client system.
申请公布号	US9176986(B2)	申请公布日期	2015.11.03
申请号	US201113309461	申请日期	2011.12.01
申请人	Google Inc.	发明人	Petrou David;Popat Ashok C.;Casey Matthew R.
分类号	G06K9/18;G06F17/30;G06K9/00;G06K9/03;G06K9/72	主分类号	G06K9/18
代理机构	Fish & Richardson P.C.	代理人	Fish & Richardson P.C.
主权项	1. A computer-implemented method of processing a visual query, performed by a server system having one or more processors and memory storing one or more programs for execution by the one or more processors, the method comprising: receiving a visual query from a client system distinct from the server system; performing optical character recognition (OCR) on the visual query to produce text recognition data representing textual characters, including a plurality of textual characters in a contiguous region of the visual query; scoring each textual character in the plurality of textual characters, wherein the scoring of a respective textual character is based on both an OCR quality score of the respective textual character alone and an OCR quality score of one or more neighboring textual characters; identifying, in accordance with the scoring, one or more high quality textual strings, each comprising a plurality of high quality textual characters from among the plurality of textual characters in the contiguous region of the visual query; retrieving a canonical document having the one or more high quality textual strings; generating a combination of the visual query and at least a portion of the canonical document; and sending the combination to the client system.
地址	Mountain View CA US