PERFORMING OPTICAL CHARACTER RECOGNITION USING SPATIAL INFORMATION OF REGIONS WITHIN A STRUCTURED DOCUMENT

Patent №

US 10,013,643

Granted

2018-07-03

Filed 2016

Owner

INTUIT INC.

Lab

AI components

4

ml · nlp · vision · hardware

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

15219888

Techniques are disclosed for facilitating optical character recognition (OCR) by identifying one or more regions in an electronic document to perform the OCR. For example a method for identifying information in an electronic document includes obtaining a set of training documents for each template of a plurality of templates for the electronic document, extracting spatial attributes for at least a first label region and at least a first corresponding value region from the set, and training a classifier model based on the extracted spatial attributes, wherein the classifier model is used to identify the information in the electronic document. The spatial attributes represent a position of at least the first label region and at least the first value region within the electronic document.

Machine learningNatural languageVisionAI hardwareG06V 30/19147G06F 18/214G06F 18/24G06Q 40/123G06T 7/11G06V 30/412G06V 30/414G06T 2207/20081+2 more

AI classification

Vision1.00
Machine learning1.00
Natural language1.00
AI hardware0.85
Planning0.46
Knowledge representation0.37
Speech0.00
Evolutionary computation0.00

Ownership

INTUIT INC.

assignment · 392610042

Assignors

YELLAPRAGADA, VIJAY, CHIANG, PEIJUN, MADDIKA, SREENEEL K.

On an employer assignment, the assignors are typically the inventors.

© 2026 NYSGPT2525 LLC