AUGMENTING VOICE SAMPLES BASED ON DISTRIBUTIONS OF SPEAKER CLASSES

Patent №

US 12,340,793

Granted

2025-06-24

Filed 2022

Owner

Interactions LLC

Lab

AI components

0

Assignment

None on record

Dataset

AIPD

Application

17851902

A method of augmenting a training dataset of voice samples is provided. An audio processing system obtains voice samples and groups the voice samples into classes of spectral representations. The system obtains warp distributions associated with the classes of spectral representations and determines spectral change ratios based on a comparison of the warp distributions. The system determines transformations based at least in part on the spectral change ratios and applies the transformations to the voice samples grouped into the classes of spectral representations to generate a set of augmented voice samples. The system compiles the training dataset using at least the set of augmented voice samples. A recognition model is trained using the training dataset.

G10L 15/063G10L 15/065G10L 15/07G10L 15/08G10L 15/10G10L 15/12G10L 21/007G10L 25/18+1 more

Ownership

Interactions LLC

From the same owner

© 2026 NYSGPT2525 LLC