Title :
Enhancing data migration performance via parallel data compression
Author :
Jonghyun Lee ; Winslett, M. ; Xiaosong Ma ; Shengke Yu
Author_Institution :
Dept. of Comput. Sci., Illinois Univ., Urbana, IL, USA
Abstract :
Scientific simulations often produce large volumes of output that are moved to another platform for visualization or storage. This long-distance migration is slow due to the data size and slow network. Compression can improve migration performance by reducing the data size, but compression is computation-intensive and so can raise costs. In this work, we show how to reduce data migration cost by incorporating compression into migration. We analyze eight scientific data sets, and propose three approaches for parallel compression of scientific data. Our results show that with reasonably fast processors and typical parallel configurations, the compression cost for large scientific data is outweighed by the performance gain obtained by migrating less data. We found that a client-side compression approach (CC) can improve I/O and migration performance by an order of magnitude. In our experiments, CC always matches or outperforms migration without compression when we overlap migration with computation, even for not very compressible dense floating point data. We also present a variant of CC that is well suited for use with implementations of two-phase I/O.
Keywords :
data analysis; data compression; natural sciences computing; parallel processing; performance evaluation; client-side compression approach; data migration performance; data size reduction; floating point data; input output performance; parallel configurations; parallel data compression; performance gain; scientific data set analysis; scientific simulations; Analytical models; Bandwidth; Compression algorithms; Computational modeling; Costs; Data compression; Data visualization; Internet; Supercomputers; Workstations;
Conference_Titel :
Parallel and Distributed Processing Symposium., Proceedings International, IPDPS 2002, Abstracts and CD-ROM
Conference_Location :
Ft. Lauderdale, FL
Print_ISBN :
0-7695-1573-8
DOI :
10.1109/IPDPS.2002.1015528