METHOD FOR DISTRIBUTED TYPE TRAINING ADAPTATION AND APPARATUS IN DEEP LEARNING FRAMEWORK AND AI ACCELERATOR CARD
Patent №
US 11,714,995
Granted
2023-08-01
Filed 2022
Owner
ZHEJIANG LAB
Lab
—
AI components
2
ml · hardware
Assignment
Recorded
Dataset
AIPD
2023_r1 edition
Application
17739205
Disclosed is a method for distributed type training adaptation and apparatus in a deep learning framework and an AI accelerator card. The method includes the following steps: S1: the deep learning framework supports single-card configuration in a newly added AI accelerator card, and sub-steps thereof are as follows: S11: the deep learning framework supports new hardware; S12: the deep learning framework supports a device thread of the new hardware; S13: the deep learning framework supports a memory operation of the new hardware; and S14: the deep learning framework supports an operator kernel function of the new hardware; S2: the deep learning framework supports multi-card configuration in the newly added AI accelerator card; S3: the deep learning framework supports tensor segmentation and multi-card distribution; and S4: the deep learning framework supports multi-card collective communication in the newly added AI accelerator card.
AI classification
Ownership
ZHEJIANG LAB
assignment · 599920931
Assignors
WANG, HONGSHENG, BAO, HUJUN, HUA, WEI, JIA, WEIQIANG
On an employer assignment, the assignors are typically the inventors.