VOICE PROCESSING METHOD, APPARATUS, DEVICE AND STORAGE MEDIUM

Patent №

US 10,839,820

Granted

2020-11-17

Filed 2018

Owner

BAIDU ONLINE NETWORK TECHNOLOGY (BEIJING) CO., LTD.

Lab

AI components

4

ml · vision · speech · hardware

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

16236261

The present application provides a voice processing method, an apparatus, a device, and a storage medium, including: acquiring a first acoustic feature of each of N voice frames, where N is a positive integer greater than 1; applying a neural network algorithm to N first acoustic features to obtain a first mask; modifying the first mask according to VAD information of the N voice frames to obtain a second mask; and processing the N first acoustic features according to the second mask to obtain a second acoustic feature, resulting in more effective noise suppression and a lower damage to the voice.

Machine learningVisionSpeechAI hardwareG10L 21/0208G06N 3/08G06N 3/09G10L 21/0272G10L 25/30G10L 25/84G10L 2021/02087

AI classification

Speech1.00
AI hardware1.00
Machine learning1.00
Vision0.97
Natural language0.41
Planning0.10
Evolutionary computation0.03
Knowledge representation0.01

Ownership

BAIDU ONLINE NETWORK TECHNOLOGY (BEIJING) CO., LTD.

assignment · 478720420

Assignors

LI, CHAO, ZHU, WEIXIN

On an employer assignment, the assignors are typically the inventors.

© 2026 NYSGPT2525 LLC