DISTRIBUTED TOKENIZATION USING SEVERAL SUBSTITUTION STEPS

Patent №

US 9,639,716

Granted

2017-05-02

Filed 2015

Owner

PROTEGRITY CORPORATION

Lab

AI components

1

nlp

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

14942668

A method for distributed tokenization of sensitive strings of characters, such as social security numbers, credit card numbers and the like, in a local server is disclosed. The method comprises the steps of receiving from a central server at least one, and preferably at least two, static token lookup tables, and receiving a sensitive string of characters. In a first tokenization step, a first substring of characters is substituted with a corresponding first token from the token lookup table(s) to form a first tokenized string of characters, wherein the first substring of characters is a substring of the sensitive string of characters. Thereafter, in a second step of tokenization, a second substring of characters is substituted with a corresponding second token from the token lookup table(s) to form a second tokenized string of characters, wherein the second substring of characters is a substring of the first tokenized string of characters. Optionally, one or more additional tokenization steps is/are used.

Natural languageG06F 21/6245G06F 16/90344G07F 7/084G07F 7/1008H04L 9/083H04L 9/0897H04L 63/0428H04L 2209/56

AI classification

Natural language1.00
Vision0.01
Knowledge representation0.00
Evolutionary computation0.00
Speech0.00
Machine learning0.00
Planning0.00
AI hardware0.00

Ownership

PROTEGRITY CORPORATION

assignment · 377740116

Assignors

MATTSSON, ULF

On an employer assignment, the assignors are typically the inventors.

© 2026 NYSGPT2525 LLC