DocumentCode :
3235061
Title :
Overlapping Methods of All-to-All Communication and FFT Algorithms for Torus-Connected Massively Parallel Supercomputers
Author :
Doi, Jun ; Negishi, Yasushi
Author_Institution :
IBM Res. Tokyo, Yamato, Japan
fYear :
2010
fDate :
13-19 Nov. 2010
Firstpage :
1
Lastpage :
9
Abstract :
Torus networks are commonly used for massively parallel computers, its performance often becomes the constraint on total application performance. Especially in an asymmetric torus network, network traffic along the longest axis is the performance bottleneck for all-to-all communication, so that it is important to schedule the longest-axis traffic smoothly. In this paper, we propose a new algorithm based on an indirect method for pipelining the all-to-all procedures using shared memory parallel threads, which (1) isolates the longest-axis traffic from other traffic, (2) schedules it smoothly and (3) overlaps all of the other traffic and overhead for the all-to-all communication behind the longest-axis traffic. The proposed method achieves up to 95% of the theoretical peak. We integrated the overlapped all-to-all method with parallel FFT algorithms. And local FFT calculations are also overlapped behind the longest-axis traffic. The FFT performance achieves up to 90% of the theoretical peak for the parallel 1D FFT.
Keywords :
fast Fourier transforms; parallel machines; shared memory systems; FFT algorithms; all-to-all communication; overlapping methods; shared memory parallel threads; torus-connected massively parallel supercomputers; Arrays; Bandwidth; Equations; Instruction sets; Message systems; Pipeline processing; Three dimensional displays;
fLanguage :
English
Publisher :
ieee
Conference_Titel :
High Performance Computing, Networking, Storage and Analysis (SC), 2010 International Conference for
Conference_Location :
New Orleans, LA
Print_ISBN :
978-1-4244-7557-5
Electronic_ISBN :
978-1-4244-7558-2
Type :
conf
DOI :
10.1109/SC.2010.38
Filename :
5645458
Link To Document :
بازگشت