SPEAKER-ADAPTIVE SYNTHESIZED VOICE

Patent №

US 8,744,853

Granted

2014-06-03

Filed 2011

Owner

INTERNATIONAL BUSINESS MACHINES CORPORATION

AI components

6

ml · nlp · vision · speech · kr · hardware

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

13319856

An objective is to provide a technique for accurately reproducing features of a fundamental frequency of a target-speaker's voice on the basis of only a small amount of learning data. A learning apparatus learns shift amounts from a reference source F0 pattern to a target F0 pattern of a target-speaker's voice. The learning apparatus associates a source F0 pattern of a learning text to a target F0 pattern of the same learning text by associating their peaks and troughs. For each of points on the target F0 pattern, the learning apparatus obtains shift amounts in a time-axis direction and in a frequency-axis direction from a corresponding point on the source F0 pattern in reference to a result of the association, and learns a decision tree using, as an input feature vector, linguistic information obtained by parsing the learning text, and using, as an output feature vector, the calculated shift amounts.

AI classification

Natural language1.00
Speech1.00
Machine learning1.00
AI hardware1.00
Vision0.96
Knowledge representation0.62
Planning0.01
Evolutionary computation0.00

Ownership

INTERNATIONAL BUSINESS MACHINES CORPORATION

assignment · 272080416

Assignors

NISHIMURA, MASAFUMI, TACHIBANA, RYUKI

On an employer assignment, the assignors are typically the inventors.

From the same owner

© 2026 NYSGPT2525 LLC