PERFORMING OPTICAL CHARACTER RECOGNITION USING SPATIAL INFORMATION OF REGIONS WITHIN A STRUCTURED DOCUMENT
Patent №
US 10,013,643
Granted
2018-07-03
Filed 2016
Owner
INTUIT INC.
Lab
—
AI components
4
ml · nlp · vision · hardware
Assignment
Recorded
Dataset
AIPD
2023_r1 edition
Application
15219888
Techniques are disclosed for facilitating optical character recognition (OCR) by identifying one or more regions in an electronic document to perform the OCR. For example a method for identifying information in an electronic document includes obtaining a set of training documents for each template of a plurality of templates for the electronic document, extracting spatial attributes for at least a first label region and at least a first corresponding value region from the set, and training a classifier model based on the extracted spatial attributes, wherein the classifier model is used to identify the information in the electronic document. The spatial attributes represent a position of at least the first label region and at least the first value region within the electronic document.
AI classification
Ownership
INTUIT INC.
assignment · 392610042
Assignors
YELLAPRAGADA, VIJAY, CHIANG, PEIJUN, MADDIKA, SREENEEL K.
On an employer assignment, the assignors are typically the inventors.