METHOD AND APPARATUS FOR TRAINING A PROSODY STATISTIC MODEL AND PROSODY PARSING, METHOD AND SYSTEM FOR TEXT TO SPEECH SYNTHESIS

Patent №

US 8,024,174

Granted

2011-09-20

Filed 2006

Owner

KABUSHIKI KAISHA TOSHIBA

Lab

AI components

4

ml · nlp · vision · speech

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

11539434

The present invention provides a method and apparatus for training a prosody statistic model and prosody parsing, a method and system for text to speech synthesis. Said method for training a prosody statistic model with a raw corpus that includes a plurality of sentences with punctuation, comprising: transforming said plurality of sentences in said raw corpus into a plurality of token sequences respectively; counting a frequency for each adjacent token pair occurring in said plurality of token sequences and frequencies of punctuation that represents a pause occurring at associated positions of said each token pair; calculating pause probabilities at said associated positions of said each token pair; and constructing said prosody statistic model based on said token pairs and said pause probabilities at associated positions thereof. With the present invention a prosody statistic model can be trained from a raw corpus without manually prosody parsing tags. And the prosody statistic model can be used in the prosody parsing and further voice synthesis.

AI classification

Natural language1.00
Speech1.00
Machine learning1.00
Vision0.81
AI hardware0.17
Knowledge representation0.10
Evolutionary computation0.00
Planning0.00

Ownership

KABUSHIKI KAISHA TOSHIBA

assignment · 188970521

Assignors

WANG, HAIFENG, LI, GUOHUA

On an employer assignment, the assignors are typically the inventors.

From the same owner

© 2026 NYSGPT2525 LLC