Efficient and portable Winograd convolutions for multi-core processors

Autor: Manuel F. Dolz, Héctor Martínez, Adrián Castelló, Pedro Alonso-Jordá, Enrique S. Quintana-Ortí
Rok vydání: 2023
Předmět:
Zdroj: The Journal of Supercomputing. 79:10589-10610
ISSN: 1573-0484
0920-8542
DOI: 10.1007/s11227-023-05088-4
Popis: We take a step forward towards developing high-performance codes for the convolution operator, based on the Winograd algorithm, that are easy to customise for general-purpose processor architectures. In our approach, augmenting the portability of the solution is achieved via the introduction of vector instructions from Intel SSE/AVX2/AVX512 and ARM NEON/SVE to exploit the single-instruction multiple-data capabilities of current processors as well as OpenMP pragmas to exploit multi-threaded parallelism. While this comes at the cost of sacrificing a fraction of the computational performance, our experimental results on three distinct processors, with Intel Xeon Skylake, ARM Cortex A57 and Fujitsu A64FX processors, show that the impact is affordable and still renders a Winograd-based solution that is competitive when compared with the lowering gemm-based convolution.
Databáze: OpenAIRE
Nepřihlášeným uživatelům se plný text nezobrazuje