portable parallel programming Latest Research Papers

AbstractIn order to adapt parallel computers to general convenient tools for computational scientists, a high-level and easy-to-use portable parallel programming paradigm is mandatory. XcalableMP, which is proposed by the XcalableMP Specification Working Group, is a directive-based language extension for Fortran and C to easily describe parallelization in programs for distributed memory parallel computers. The Omni XcalableMP compiler, which is provided as a reference XcalableMP compiler, is currently implemented as a source-to-source translator. It converts XcalableMP programs to standard MPI programs, which can be easily compiled by the native Fortran compiler and executed on most of parallel computers. A three-dimensional Eulerian fluid code written in Fortran is parallelized by XcalableMP using two different programming models with the ordinary domain decomposition method, and its performances are measured on the K computer. Programs converted by the Omni XcalableMP compiler prevent native Fortran compiler optimizations and show lower performance than that of hand-coded MPI programs. Finally almost the same performances are obtained by using specific compiler options of the native Fortran compiler in the case of a global-view programming model, but performance degradation is not improved by specifying any native compiler options when the code is parallelized by a local-view programming model.

Download Full-text

Towards Implicit Memory Management for Portable Parallel Programming in C++

Proceedings of the 2020 Asia Service Sciences and Software Engineering Conference ◽

10.1145/3399871.3399881 ◽

2020 ◽

Author(s):

Vladyslav Kucher ◽

Sergei Gorlatch

Keyword(s):

Implicit Memory ◽

Parallel Programming ◽

Memory Management ◽

Portable Parallel Programming

Download Full-text

Performance portable parallel programming of heterogeneous stencils across shared-memory platforms with modern Intel processors

The International Journal of High Performance Computing Applications ◽

10.1177/1094342019828153 ◽

2019 ◽

Vol 33 (3) ◽

pp. 534-553 ◽

Cited By ~ 4

Author(s):

Lukasz Szustak ◽

Pawel Bratek

Keyword(s):

Shared Memory ◽

Memory Systems ◽

Optimization Techniques ◽

Utilization Rate ◽

Problem Size ◽

Geophysical Model ◽

Step Procedure ◽

Wide Range ◽

Sustained Performance ◽

Portable Parallel Programming

In this work, we take up the challenge of performance portable programming of heterogeneous stencil computations across a wide range of modern shared-memory systems. An important example of such computations is the Multidimensional Positive Definite Advection Transport Algorithm (MPDATA), the second major part of the dynamic core of the EULAG geophysical model. For this aim, we develop a set of parametric optimization techniques and four-step procedure for customization of the MPDATA code. Among these techniques are: islands-of-cores strategy, (3+1)D decomposition, exploiting data parallelism and simultaneous multithreading, data flow synchronization, and vectorization. The proposed adaptation methodology helps us to develop the automatic transformation of the MPDATA code to achieve high sustained scalable performance for all tested ccNUMA platforms with Intel processors of last generations. This means that for a given platform, the sustained performance of the new code is kept at a similar level, independently of the problem size. The highest performance utilization rate of about 41–46% of the theoretical peak, measured for all benchmarks, is provided for any of the two-socket servers based on Skylake-SP (SKL-SP), Broadwell, and Haswell CPU architectures. At the same time, the four-socket server with SKL-SP processors achieves the highest sustained performance of around 1.0–1.1 Tflop/s that corresponds to about 33% of the peak.

Download Full-text

CUDAGA: A Portable Parallel Programming Model for GPU Cluster

Cloud Computing and Big Data - Lecture Notes in Computer Science ◽

10.1007/978-3-319-28430-9_16 ◽

2015 ◽

pp. 203-216

Author(s):

Yong Chen ◽

Hai Jin ◽

Dechao Xu ◽

Ran Zheng ◽

Haocheng Liu ◽

...

Keyword(s):

Parallel Programming ◽

Programming Model ◽

Gpu Cluster ◽

Parallel Programming Model ◽

Portable Parallel Programming

Download Full-text

Portable Parallel Programming on Cloud and HPC: Scientific Applications of Twister4Azure

2011 Fourth IEEE International Conference on Utility and Cloud Computing ◽

10.1109/ucc.2011.23 ◽

2011 ◽

Cited By ~ 6

Author(s):

T. Gunarathne ◽

Bingjing Zhang ◽

Tak-Lon Wu ◽

J. Qiu

Keyword(s):

Parallel Programming ◽

Scientific Applications ◽

Portable Parallel Programming

Download Full-text

The performance analysis of portable parallel programming interface MpC for SDSM and pthread

CCGrid 2005. IEEE International Symposium on Cluster Computing and the Grid, 2005. ◽

10.1109/ccgrid.2005.1558656 ◽

2005 ◽

Cited By ~ 1

Author(s):

H. Midorikawa

Keyword(s):

Performance Analysis ◽

Parallel Programming ◽

Programming Interface ◽

Portable Parallel Programming

Download Full-text

Portable parallel programming for the dynamic load balancing of unstructured grid applications

Proceedings 13th International Parallel Processing Symposium and 10th Symposium on Parallel and Distributed Processing. IPPS/SPDP 1999 ◽

10.1109/ipps.1999.760497 ◽

2003 ◽

Cited By ~ 6

Author(s):

R. Biswas ◽

S.K. Das ◽

D. Harvey ◽

L. Oliker

Keyword(s):

Load Balancing ◽

Parallel Programming ◽

Dynamic Load ◽

Unstructured Grid ◽

Dynamic Load Balancing ◽

Grid Applications ◽

Portable Parallel Programming

Download Full-text

An introduction to the portable parallel programming language Seymour

[1989] Proceedings of the Thirteenth Annual International Computer Software & Applications Conference ◽

10.1109/cmpsac.1989.65063 ◽

2003 ◽

Author(s):

R. Miller ◽

Q.F. Stout

Keyword(s):

Parallel Programming ◽

Programming Language ◽

Portable Parallel Programming ◽

Parallel Programming Language

Download Full-text

portable parallel programming
Recently Published Documents

TOTAL DOCUMENTS

H-INDEX

Comparing Julia to Performance Portable Parallel Programming Models for HPC

Implicit Data Layout Optimization for Portable Parallel Programming in C++

Three-Dimensional Fluid Code with XcalableMP

Towards Implicit Memory Management for Portable Parallel Programming in C++

Performance portable parallel programming of heterogeneous stencils across shared-memory platforms with modern Intel processors

CUDAGA: A Portable Parallel Programming Model for GPU Cluster

Portable Parallel Programming on Cloud and HPC: Scientific Applications of Twister4Azure

The performance analysis of portable parallel programming interface MpC for SDSM and pthread

Portable parallel programming for the dynamic load balancing of unstructured grid applications

An introduction to the portable parallel programming language Seymour

Export Citation Format

portable parallel programmingRecently Published Documents

TOTAL DOCUMENTS

H-INDEX

Comparing Julia to Performance Portable Parallel Programming Models for HPC

Implicit Data Layout Optimization for Portable Parallel Programming in C++

Three-Dimensional Fluid Code with XcalableMP

Towards Implicit Memory Management for Portable Parallel Programming in C++

Performance portable parallel programming of heterogeneous stencils across shared-memory platforms with modern Intel processors

CUDAGA: A Portable Parallel Programming Model for GPU Cluster

Portable Parallel Programming on Cloud and HPC: Scientific Applications of Twister4Azure

The performance analysis of portable parallel programming interface MpC for SDSM and pthread

Portable parallel programming for the dynamic load balancing of unstructured grid applications

An introduction to the portable parallel programming language Seymour

portable parallel programming
Recently Published Documents