OPTIMIZED AND SCALABLE SPARSE TRIANGULAR LINEAR SYSTEMS ON NETWORKS OF ACCELERATORS

Patent №

US 10,936,697

Granted

2021-03-02

Filed 2018

Owner

ADVANCED MICRO DEVICES, INC.

Lab

AI components

1

hardware

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

16044145

A method includes storing a first portion of a sparse triangular matrix in a local memory and launching a kernel for executing a set of workgroups. The first portion includes a plurality of row blocks, and each workgroup in the set of workgroups is associated with one of the plurality of row blocks. The method also includes, for each workgroup in the set of workgroups, solving the row block. The row block is solved by, for each row segment of a first subset of row segments in the row block, calculating a partial sum for the row segment based on one or more matrix elements in the row segment, and writing the partial sum to a remote memory of a first remote processing unit prior to terminating the kernel.

AI hardwareG06F 17/16G06F 9/3001G06F 9/3838G06F 9/3889G06F 17/12

AI classification

AI hardware1.00
Machine learning0.01
Planning0.00
Natural language0.00
Vision0.00
Knowledge representation0.00
Evolutionary computation0.00
Speech0.00

Ownership

ADVANCED MICRO DEVICES, INC.

assignment · 464460947

Assignors

GREATHOUSE, JOSEPH L, HAMIDOUCHE, KHALED, LEBEANE, MICHAEL W, MALAYA, NICHOLAS P

On an employer assignment, the assignors are typically the inventors.

© 2026 NYSGPT2525 LLC