Indexed metadata

High-performance implementation of the level-3 BLAS

Kazushige Goto, Robert Van De Geijn

Source record

Source: Crossref

Published: Jul 22, 2008

DOI: 10.1145/1377603.1377607

Open original source ↗

Source abstract

A simple but highly effective approach for transforming high-performance implementations on cache-based architectures of matrix-matrix multiplication into implementations of other commonly used matrix-matrix computations (the level-3 BLAS) is presented. Exceptional performance is demonstrated on various architectures.

Evidence graph

No public relationships recorded yet.

Integrity note: This page is a factual metadata record created by deterministic ingestion. It is not a claim that the work moves a mathematical frontier or has been independently verified.