METHOD AND SYSTEM FOR COLLECTING DATA FROM A PLURALITY OF MACHINE READABLE DOCUMENTS

Patent №

US 7,668,372

Granted

2010-02-23

Filed 2006

Owner

OCE DOCUMENT TECHNOLOGIES GMBH

+1 more

Lab

AI components

2

nlp · planning

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

10571239

In a method and system for collection of data from documents present in machine-readable form, at least one already-processed document stored as a template and designated as a template document is associated with a document to be processed designated as a read document. Fields for data to be extracted are defined in the template document. Data contained in the read document are already extracted from regions that correspond to the fields in the template document. Should an error have occurred or no suitable template document having been associated given the automatic extraction of the data, the read document is shown on a screen and fields are manually inputted in the read document from which the data are extracted. After the manual input of the fields in the read document, the read document with field specifications is stored as a new template document or the previous template document is corrected corresponding to the newly input fields.

Natural languagePlanningG06V 30/416G06V 30/127G06V 30/1444G06V 30/10

AI classification

Planning1.00
Natural language0.82
Vision0.06
Knowledge representation0.03
Machine learning0.01
AI hardware0.01
Evolutionary computation0.00
Speech0.00

Ownership

OCE DOCUMENT TECHNOLOGIES GMBH

assignment · 186030771

OPEN TEXT S.A.

assignment · 272920666

Assignors

SCHIEHLEN, MATTHIAS

On an employer assignment, the assignors are typically the inventors.

© 2026 NYSGPT2525 LLC