Vectorization, threading, and cache-blocking considerations for hydrocodes on emerging architectures

Fung, J.; Aulwes, R. T.; Bement, M. T.; Campbell, J. M.; Ferenbaugh, C. R.; Jean, B. A.; Kelley, T. M.; Kenamond, M. A.; Lally, B. R.; Lovegrove, E. G.; Nelson, E. M.; Powell, D. M.

doi:10.1002/fld.4063

Title: Vectorization, threading, and cache-blocking considerations for hydrocodes on emerging architectures

Journal Article · Tue Jul 14 00:00:00 EDT 2015 · International Journal for Numerical Methods in Fluids

DOI:https://doi.org/10.1002/fld.4063· OSTI ID:1214827

Fung, J. ^[1]; Aulwes, R. T. ^[1]; Bement, M. T. ^[1]; Campbell, J. M. ^[1]; Ferenbaugh, C. R. ^[1]; Jean, B. A. ^[1]; Kelley, T. M. ^[1]; Kenamond, M. A. ^[1]; Lally, B. R. ^[1]; Lovegrove, E. G. ^[2]; Nelson, E. M. ^[1]; Powell, D. M. ^[3]

Los Alamos National Lab. (LANL), Los Alamos, NM (United States)
Univ. of California, Santa Cruz, CA (United States)
Stanford Univ., CA (United States)

This work reports on considerations for improving computational performance in preparation for current and expected changes to computer architecture. The algorithms studied will include increasingly complex prototypes for radiation hydrodynamics codes, such as gradient routines and diffusion matrix assembly (e.g., in [1-6]). The meshes considered for the algorithms are structured or unstructured meshes. The considerations applied for performance improvements are meant to be general in terms of architecture (not specifically graphical processing unit (GPUs) or multi-core machines, for example) and include techniques for vectorization, threading, tiling, and cache blocking. Out of a survey of optimization techniques on applications such as diffusion and hydrodynamics, we make general recommendations with a view toward making these techniques conceptually accessible to the applications code developer. Published 2015. This article is a U.S. Government work and is in the public domain in the USA.

View Accepted Manuscript (DOE)

Cite

Export

Save

Research Organization:: Los Alamos National Lab. (LANL), Los Alamos, NM (United States)

Sponsoring Organization:: USDOE

Grant/Contract Number:: AC52-06NA25396

OSTI ID:: 1214827

Report Number(s):: LA-UR-14-21299

Journal Information:: International Journal for Numerical Methods in Fluids, Vol. 79, Issue 11; ISSN 0271-2091

Publisher:: WileyCopyright Statement

Country of Publication:: United States

Language:: English

Similar Records

Efficient Machine Learning Approach for Optimizing Scientific Computing Applications on Emerging HPC Architectures

Thesis/Dissertation · Mon May 01 00:00:00 EDT 2017 · OSTI ID:1214827

Arumugam, Kamesh

Deploy threading in Nalu solver stack

Technical Report · Mon Oct 01 00:00:00 EDT 2018 · OSTI ID:1214827

Prokopenko, Andrey; Thomas, Stephen; Swirydowicz, Kasia; +4 more

Data Locality Enhancement of Dynamic Simulations for Exascale Computing (Final Report)

Technical Report · Fri Nov 29 00:00:00 EST 2019 · OSTI ID:1214827

Shen, Xipeng

Related Subjects

97 MATHEMATICS AND COMPUTING
Lagrangian Hydrodynamics
Arbitrary Lagrangian Eulerian (ALE) Methods
Radiation Hydrodynamics
Computer Science and Advanced Architectures

Title: Vectorization, threading, and cache-blocking considerations for hydrocodes on emerging architectures

Citation Formats

Similar Records

Related Subjects