Indexed metadata
High-performance implementation of the level-3 BLAS
Kazushige Goto, Robert Van De Geijn
Source record
Source: Crossref
Published: Jul 22, 2008
DOI: 10.1145/1377603.1377607
Open original source ↗Source abstract
A simple but highly effective approach for transforming high-performance implementations on cache-based architectures of matrix-matrix multiplication into implementations of other commonly used matrix-matrix computations (the level-3 BLAS) is presented. Exceptional performance is demonstrated on various architectures.
Evidence graph
No public relationships recorded yet.
Integrity note: This page is a factual metadata record created by deterministic ingestion. It is not a claim that the work moves a mathematical frontier or has been independently verified.