skip to main content
OSTI.GOV title logo U.S. Department of Energy
Office of Scientific and Technical Information

Title: Lattice Boltzmann Simulation Optimization on Leading Multicore Platforms

Conference ·
OSTI ID:964372

We present an auto-tuning approach to optimize application performance on emerging multicore architectures. The methodology extends the idea of search-based performance optimizations, popular in linear algebra and FFT libraries, to application-specific computational kernels. Our work applies this strategy to a lattice Boltzmann application (LBMHD) that historically has made poor use of scalar microprocessors due to its complex data structures and memory access patterns. We explore one of the broadest sets of multicore architectures in the HPC literature, including the Intel Clovertown, AMD Opteron X2, Sun Niagara2, STI Cell, as well as the single core Intel Itanium2. Rather than hand-tuning LBMHD for each system, we develop a code generator that allows us identify a highly optimized version for each platform, while amortizing the human programming effort. Results show that our auto-tuned LBMHD application achieves up to a 14x improvement compared with the original code. Additionally, we present detailed analysis of each optimization, which reveal surprising hardware bottlenecks and software challenges for future multicore systems and applications.

Research Organization:
Lawrence Berkeley National Lab. (LBNL), Berkeley, CA (United States)
Sponsoring Organization:
Computational Research Division
DOE Contract Number:
DE-AC02-05CH11231
OSTI ID:
964372
Report Number(s):
LBNL-2165E; TRN: US200919%%221
Resource Relation:
Conference: International Parallel and Distributed Processing Symposium (IPDPS), 2008, Miami, FL, 04/14-28/2008
Country of Publication:
United States
Language:
English

Similar Records

Lattice Boltzmann simulation optimization on leading multicore platforms
Conference · Tue Jan 01 00:00:00 EST 2008 · OSTI ID:964372

Optimization of a Lattice Boltzmann Computation on State-of-the-Art Multicore Platforms
Journal Article · Fri Apr 10 00:00:00 EDT 2009 · Journal of Parallel and Distributed Computing · OSTI ID:964372

PERI - Auto-tuning Memory Intensive Kernels for Multicore
Conference · Tue Jun 24 00:00:00 EDT 2008 · OSTI ID:964372