Porting mpi4py-fft to GPU

Baumann, Thomas; Speck, Robert

Conference Presentation (After Call)

FZJ-2024-04850

Porting mpi4py-fft to GPU

Baumann, T. (Corresponding author)FZJ* ; Speck, R.FZJ*

2024

16th JLESC Workshop, JLESC16, Kobe, Japan, 16 Apr 2024 - 18 Apr 2024 [10.34734/FZJ-2024-04850]

This record in other databases:

Please use a persistent id in citations: doi:10.34734/FZJ-2024-04850

Abstract: The mpi4py-fft library enables distributed fast Fourier transforms on CPUs with an easy to use interface and scales very well. We attempt to port this to GPUs, which significantly outperform the CPU counterpart at a given node count. While the porting is straightforward for the most part, the best communication strategy is still an open question for us.The algorithm relies on MPI alltoallw. Even with CUDA-aware MPI, this exhibits very poor performance on the Juelich computers. By replacing it with a custom communication strategy, throughput can be increased at a slight loss of generality. We would like to discuss optimising the strategy, or even if the performance of alltoallw can be increased by some measure.

Contributing Institute(s):

Jülich Supercomputing Center (JSC)

Research Program(s):

Appears in the scientific report 2024

Database coverage:
OpenAccess

Click to display QR Code for this record

The record appears in these collections:
Dokumenttypen > Präsentationen > Konferenzvorträge
Workflowsammlungen > Öffentliche Einträge
Institutssammlungen > JSC
Publikationsdatenbank
Open Access

Datensatz erzeugt am 2024-07-12, letzte Änderung am 2024-12-18

Ähnliche Datensätze

OpenAccess:

PDF

Dieses Dokument bewerten:

(Bisher nicht rezensiert)

Zum persönlichen Korb hinzufügen
Export als Author List with IDs BibTeX (UTF-8), EndNote XML, EndNote Text, RIS, MARC, Print MARC, MARCXML, DC,
Request correction
Submit fulltext

Gast :: Anmelden JuSER
		Suchen		Absenden		Personalisieren Ihre Benachrichtigungen Ihre Körbe Ihre Suchanfragen		Hilfe