Optimizing Noncontiguous Accesses in MPI-IO

Computer Science – Distributed – Parallel – and Cluster Computing

Scientific paper

Rate now

[ 0.00 ] – not rated yet Voters 0 Comments 0

Details Optimizing Noncontiguous Accesses in MPI-IO Optimizing Noncontiguous Accesses in MPI-IO

: 2003-10-15
: arxiv.org/abs/cs/0310029v1
: Parallel Computing 28(1) (January 2002), pp. 83-105
: Computer Science
: Distributed, Parallel, and Cluster Computing

: 18 pages, 12 figures
: Scientific paper
: The I/O access patterns of many parallel applications consist of accesses to a large number of small, noncontiguous pieces of data. If an application's I/O needs are met by making many small, distinct I/O requests, however, the I/O performance degrades drastically. To avoid this problem, MPI-IO allows users to access noncontiguous data with a single I/O function call, unlike in Unix I/O. In this paper, we explain how critical this feature of MPI-IO is for high performance and how it enables implementations to perform optimizations. We first provide a classification of the different ways of expressing an application's I/O needs in MPI-IO--we classify them into four levels, called level~0 through level~3. We demonstrate that, for applications with noncontiguous access patterns, the I/O performance improves dramatically if users write their applications to make level-3 requests (noncontiguous, collective) rather than level-0 requests (Unix style). We then describe how our MPI-IO implementation, ROMIO, delivers high performance for noncontiguous requests. We explain in detail the two key optimizations ROMIO performs: data sieving for noncontiguous requests from one process and collective I/O for noncontiguous requests from multiple processes. We describe how we have implemented these optimizations portably on multiple machines and file systems, controlled their memory requirements, and also achieved high performance. We demonstrate the performance and portability with performance results for three applications--an astrophysics-application template (DIST3D), the NAS BTIO benchmark, and an unstructured code (UNSTRUC)--on five different parallel machines: HP Exemplar, IBM SP, Intel Paragon, NEC SX-4, and SGI Origin2000.

Affiliated with

Gropp William

Computer Science – Distributed – Parallel – and Cluster Computing

Scientist

[ 0.00 ] – not rated yet Voters 0 Comments 0

Lusk Ewing

Computer Science – Distributed – Parallel – and Cluster Computing

Scientist

[ 0.00 ] – not rated yet Voters 0 Comments 0

Thakur Rajeev

Computer Science – Distributed – Parallel – and Cluster Computing

Scientist

[ 0.00 ] – not rated yet Voters 0 Comments 0

Also associated with

No associations

LandOfFree

Say what you really think

Search LandOfFree.com for scientists and scientific papers. Rate them and share your experience with other people.

Rating

Optimizing Noncontiguous Accesses in MPI-IO does not yet have a rating. At this time, there are no reviews or comments for this scientific paper.
If you have personal experience with Optimizing Noncontiguous Accesses in MPI-IO, we encourage you to share that experience with our LandOfFree.com community. Your opinion is very important and Optimizing Noncontiguous Accesses in MPI-IO will most certainly appreciate the feedback.

Rate now

Comments { 0 }

Profile ID: LFWR-SCP-O-527507

All data on this website is collected from public sources. Our data reflects the most accurate information available at the time of publication.

Canada

Charities
Companies
MP Candidates
Patents
Employee Salary Disclosure

World

Places of the World
Scientific Papers

United States

Banks
Companies
Counties
Patents
Employee Salary Disclosure