This paper examines common implementations of linear algebra algorithms, such as matrix-vector multiplication, matrix-matrix multiplication and the solution of linear equations. The different versions are examined for efficiency on a computer architecture which uses vector processing and has pipelined instruction execution. By using the advanced architectural features of such machines, one can usually achieve maximum performance, and tremendous improvements in terms of execution speed can be seen over conventional computers.
No takes yet. Share an insight, caveat, or question.
Dongarra et al. (1984) studied this question.
Synapse has enriched 3 closely related papers on similar clinical questions. Consider them for comparative context: