• DocumentCode
    1415093
  • Title

    Data locality exploitation in the decomposition of regular domain problems

  • Author

    Prieto, Manuel ; Llorente, Ignacio M. ; Tirado, Franscisco

  • Author_Institution
    Dept. de Arquitectura de Comput. y Autom., Univ. Complutense de Madrid, Spain
  • Volume
    11
  • Issue
    11
  • fYear
    2000
  • fDate
    11/1/2000 12:00:00 AM
  • Firstpage
    1141
  • Lastpage
    1150
  • Abstract
    The aim of this paper is to study the effect of local memory hierarchy and communication network exploitation on message sending and the influence of this effect on the decomposition of regular applications. In particular, we have considered two different parallel computers, a Cray T3E-900 and an SGI Origin 2000. In both systems, the bandwidth reduction due to non-unit-stride memory access is quite significant and could be more important than the reduction due to contention in the network. These conclusions affect the choice of optimal decompositions for regular domains problems. Thus, although traditional 3D decompositions lead to lower inherent communication-to-computation ratios and could exploit more efficiently the interconnection network, lower dimensional decompositions are found to be more efficient due to the data decomposition effects on the spatial locality of the messages to be communicated. This increasing importance of local optimisations has also been shown using a well-known communication-computation overlapping technique which increases execution time, instead of reducing it as we could expect, due to poor cache memory exploitation.
  • Keywords
    message passing; parallel processing; performance evaluation; Cray T3E-900; SGI Origin 2000; cache memory exploitation; communication network exploitation; communication-computation overlapping; data locality exploitation; local memory hierarchy; message sending; parallel computers; regular domain problems decomposition; spatial locality; Application software; Bandwidth; Cache memory; Communication networks; Concurrent computing; Costs; Intelligent networks; Iterative algorithms; Microprocessors; Multiprocessor interconnection networks;
  • fLanguage
    English
  • Journal_Title
    Parallel and Distributed Systems, IEEE Transactions on
  • Publisher
    ieee
  • ISSN
    1045-9219
  • Type

    jour

  • DOI
    10.1109/71.888635
  • Filename
    888635