Show simple item record

contributor authorJohn C. Pau
contributor authorBrett F. Sanders
date accessioned2017-05-08T21:13:15Z
date available2017-05-08T21:13:15Z
date copyrightMarch 2006
date issued2006
identifier other%28asce%290887-3801%282006%2920%3A2%2899%29.pdf
identifier urihttp://yetl.yabesh.ir/yetl/handle/yetl/43260
description abstractExplicit total variation diminishing finite-volume schemes are being adopted on a widespread basis for the solution of depth-averaged hydrodynamic equations. Explicit schemes are constrained by the Courant-Friedrichs-Lewy condition for stability purposes, and therefore require use of a small time step. As grid resolution increases, the ratio of run time to integration time may approach unity, so strategies to reduce run times are sought. This paper characterizes the performance gains of two parallel computing optimizations exercised on two different computer architectures. The optimizations include removal of explicit synchronization mechanisms (Level 1) and conversion of blocking to nonblocking communications (Level 2). Our findings show that Level 1 always improves speed-up over Level 0, while the effectiveness of Level 2 over Level 1 is mixed. Level 2 results in the best performance on a system with a relatively small bandwidth (100 Mb) interconnect switch, but in a few cases involving a system with a gigabit interconnect switch, Level 2 actually leads to slow-down as compared to Level 1. Level 2 was found to be more difficult to implement than Level 1, and the resulting code was less modular and more difficult to read. Overall, the marginal performance improvements of nonblocking communications (Level 2) cannot justify the effort to realize the optimization and the cost of a less readable program. In the context of algorithm development, we emphasize delaying optimizations until a correct parallel implementation has been obtained. The benefit is that optimizations best suited to the underlying hardware architecture can be identified.
publisherAmerican Society of Civil Engineers
titlePerformance of Parallel Implementations of an Explicit Finite-Volume Shallow-Water Model
typeJournal Paper
journal volume20
journal issue2
journal titleJournal of Computing in Civil Engineering
identifier doi10.1061/(ASCE)0887-3801(2006)20:2(99)
treeJournal of Computing in Civil Engineering:;2006:;Volume ( 020 ):;issue: 002
contenttypeFulltext


Files in this item

Thumbnail

This item appears in the following Collection(s)

Show simple item record