PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
February 19, 2024The Journal of Supercomputing5 citationsOpen Access

Parallel GEMM-based convolution for deep learning on multicore RISC-V processors

View Full Paper
CRCristián RamírezACAdrián CastellóHMHéctor Martínez

Key Points

Key points are not available for this paper at this time.

Abstract

Abstract We address the efficient implementation of the convolution operator on the GAP8 parallel ultra-low power platform (PULP), a heterogeneous multi-core processor equipped with a fabric controller (FC); a cluster of eight compute cores; and a four-level memory hierarchy with scratchpads instead of conventional, hardware-assisted cache memories. Our solution for this platform transforms the convolution into a general matrix–matrix multiplication ( gemm ) via the lowering approach, demonstrating that it is possible to attain reasonable performance on the GAP8 by carefully adapting techniques such as tiling and loop parallelism, which are mainstream in the multi-threaded, cache-aware realization of gemm .

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Ramírez et al. (2024) studied this question.

synapsesocial.com/papers/68e78968b6db6435876fbe3chttps://doi.org/10.1007/s11227-024-05927-y
Ask AI
Helpful
Bookmark
Share
View Full Paper