2013/10/30 by Alexandre da Costa Sena, Alexandre Sena, Aline Nascimento +9
Computer Science · #Distributed #Distributed and Parallel Computing Systems #FOS: Computer and information sciences #Neural Networks and Applications #Parallel #Parallel Computing and Optimization Techniques #and Cluster Computing (cs.DC) #cs.DC
paper · pdf · doi:10.48550/arxiv.1310.8232
arxiv created 2013/10/30 · openalex publication_date 2013/10/30 · arxiv updated 2013/10/31 · openalex created_date 2016/06/24 · openalex updated_date 2026/07/28
Although modern supercomputers are composed of multicore machines, one can find scientists that still execute their legacy applications which were developed to monocore cluster where memory hierarchy is dedicated to a sole core. The main objective of this paper is to propose and evaluate an algorithm that identify an efficient blocksize to be applied on MPI stencil computations on multicore machines. Under the light of an extensive experimental analysis, this work shows the benefits of identifying blocksizes that will dividing data on the various cores and suggest a methodology that explore the memory hierarchy available in modern machines.