Identifying matching canonical documents in response to a visual query and in accordance with geographic information
First Claim
1. A computer-implemented method of processing a visual query performed by a server system having one or more processors and memory storing one or more programs for execution by the one or more processors, the method comprising:
- at the server system;
receiving from a client system distinct from the server system a visual query and information identifying a geographic location of the client system;
performing optical character recognition (OCR) on the visual query to produce text recognition data representing textual characters, including a plurality of textual characters in a contiguous region of the visual query;
scoring each textual character in the plurality of textual characters, including scoring each textual character in the plurality of textual characters in accordance with the geographic location of the client system, wherein the scoring of a respective textual character comprises generating a language-conditional character likelihood for the respective textual character indicating how likely the respective textual character and a set of characters that precede the respective textual character in a text segment concord with a language model selected in accordance with the geographic location of the client system;
identifying, in accordance with the scoring, one or more high quality textual strings, each comprising a plurality of high quality textual characters from among the plurality of textual characters in the contiguous region of the visual query;
retrieving a canonical document having the one or more high quality textual strings; and
sending at least a portion of the canonical document to the client system.
2 Assignments
0 Petitions
Accused Products
Abstract
A server system receives a visual query from a client system distinct from the server system. The server system performs optical character recognition (OCR) on the visual query to produce text recognition data representing textual characters, including a plurality of textual characters in a contiguous region of the visual query. The server system scores each textual character in the plurality of textual characters in accordance with the geographic location of the client system. The server system identifies, in accordance with the scoring, one or more high quality textual strings, each comprising a plurality of high quality textual characters from among the plurality of textual characters in the contiguous region of the visual query. Then the server system retrieves a canonical document having the one or more high quality textual strings and sends at least a portion of the canonical document to the client system.
72 Citations
20 Claims
-
1. A computer-implemented method of processing a visual query performed by a server system having one or more processors and memory storing one or more programs for execution by the one or more processors, the method comprising:
at the server system; receiving from a client system distinct from the server system a visual query and information identifying a geographic location of the client system; performing optical character recognition (OCR) on the visual query to produce text recognition data representing textual characters, including a plurality of textual characters in a contiguous region of the visual query; scoring each textual character in the plurality of textual characters, including scoring each textual character in the plurality of textual characters in accordance with the geographic location of the client system, wherein the scoring of a respective textual character comprises generating a language-conditional character likelihood for the respective textual character indicating how likely the respective textual character and a set of characters that precede the respective textual character in a text segment concord with a language model selected in accordance with the geographic location of the client system; identifying, in accordance with the scoring, one or more high quality textual strings, each comprising a plurality of high quality textual characters from among the plurality of textual characters in the contiguous region of the visual query; retrieving a canonical document having the one or more high quality textual strings; and sending at least a portion of the canonical document to the client system. - View Dependent Claims (2, 3, 4, 5, 6, 7, 8, 9, 10, 18)
-
11. A method of processing a visual query performed by a server system having one or more processors and memory storing one or more programs for execution by the one or more processors, the method comprising:
at the server system; receiving from a client system distinct from the server system a visual query and information identifying a geographic location of the client system; performing optical character recognition (OCR) on the visual query to produce text recognition data representing textual characters, including a plurality of textual characters in a contiguous region of the visual query; scoring each textual character in the plurality of textual characters, including scoring each textual character in the plurality of textual characters in accordance with the geographic location of the client system; identifying, in accordance with the scoring, one or more high quality textual strings, each comprising a plurality of high quality textual characters from among the plurality of textual characters in the contiguous region of the visual query; retrieving a canonical document having the one or more high quality textual strings, the retrieving comprising; calculating a quality score corresponding to at least one respective high quality textual string of the one or more high quality textual strings; retrieving an image version of the canonical document if the quality score is below a predetermined value; and retrieving a machine readable text version of the canonical document if the quality score is at or above a predetermined value; and sending at least a portion of the canonical document to the client system.
-
12. A server system, for processing a visual query, comprising:
-
one or more central processing units for executing programs; memory storing one or more programs be executed by the one or more central processing units; the one or more programs comprising instructions for; receiving a visual query from a client system and information identifying a geographic location of the client system; performing optical character recognition (OCR) on the visual query to produce text recognition data representing textual characters, including a plurality of textual characters in a contiguous region of the visual query; scoring each textual character in the plurality of textual characters, including scoring each textual character in the plurality of textual characters in accordance with the geographic location of the client system, wherein the scoring of a respective textual character comprises generating a language-conditional character likelihood for the respective textual character indicating how likely the respective textual character and a set of characters that precede the respective textual character in a text segment concord with a language model selected in accordance with the geographic location of the client system; identifying, in accordance with the scoring, one or more high quality textual strings, each comprising a plurality of high quality textual characters from among the plurality of textual characters in the contiguous region of the visual query; retrieving a canonical document having the one or more high quality textual strings; and sending at least a portion of the canonical document to the client system. - View Dependent Claims (13, 14)
-
-
15. A non-transitory computer readable storage medium storing one or more programs configured for execution by a computer, the one or more programs comprising instructions for:
-
receiving a visual query from a client system and information identifying a geographic location of the client system; performing optical character recognition (OCR) on the visual query to produce text recognition data representing textual characters, including a plurality of textual characters in a contiguous region of the visual query; scoring each textual character in the plurality of textual characters, including scoring each textual character in the plurality of textual characters in accordance with the geographic location of the client system, wherein the scoring of a respective textual character comprises generating a language-conditional character likelihood for the respective textual character indicating how likely the respective textual character and a set of characters that precede the respective textual character in a text segment concord with a language model selected in accordance with the geographic location of the client system; identifying, in accordance with the scoring, one or more high quality textual strings, each comprising a plurality of high quality textual characters from among the plurality of textual characters in the contiguous region of the visual query; retrieving a canonical document having the one or more high quality textual strings; and sending at least a portion of the canonical document to the client system. - View Dependent Claims (16, 17)
-
-
19. A server system, for processing a visual query, comprising:
-
one or more central processing units for executing programs; memory storing one or more programs be executed by the one or more central processing units; the one or more programs comprising instructions for; receiving from a client system distinct from the server system a visual query and information identifying a geographic location of the client system; performing optical character recognition (OCR) on the visual query to produce text recognition data representing textual characters, including a plurality of textual characters in a contiguous region of the visual query; scoring each textual character in the plurality of textual characters, including scoring each textual character in the plurality of textual characters in accordance with the geographic location of the client system; identifying, in accordance with the scoring, one or more high quality textual strings, each comprising a plurality of high quality textual characters from among the plurality of textual characters in the contiguous region of the visual query; retrieving a canonical document having the one or more high quality textual strings, the retrieving comprising; calculating a quality score corresponding to at least one respective high quality textual string of the one or more high quality textual strings; retrieving an image version of the canonical document if the quality score is below a predetermined value; and retrieving a machine readable text version of the canonical document if the quality score is at or above a predetermined value; and sending at least a portion of the canonical document to the client system.
-
-
20. A non-transitory computer readable storage medium storing one or more programs configured for execution by a computer, the one or more programs comprising instructions for:
-
receiving from a client system distinct from the server system a visual query and information identifying a geographic location of the client system; performing optical character recognition (OCR) on the visual query to produce text recognition data representing textual characters, including a plurality of textual characters in a contiguous region of the visual query; scoring each textual character in the plurality of textual characters, including scoring each textual character in the plurality of textual characters in accordance with the geographic location of the client system; identifying, in accordance with the scoring, one or more high quality textual strings, each comprising a plurality of high quality textual characters from among the plurality of textual characters in the contiguous region of the visual query; retrieving a canonical document having the one or more high quality textual strings, the retrieving comprising; calculating a quality score corresponding to at least one respective high quality textual string of the one or more high quality textual strings; retrieving an image version of the canonical document if the quality score is below a predetermined value; and retrieving a machine readable text version of the canonical document if the quality score is at or above a predetermined value; and sending at least a portion of the canonical document to the client system.
-
Specification