Skip to main content

Targeting Atmospheric Simulation Algorithms for Large Distributed Memory GPU Accelerated Computers...

by Matthew R Norman
Publication Type
Book Chapter
Publication Date
Page Number
Publisher Name
Publisher Location
New York, New Jersey, United States of America

Computing platforms are increasingly moving to accelerated architectures, and here we deal particularly with GPUs. In [15], a method was developed for atmospheric simulation to improve efficiency on large distributed memory machines by reducing communication demand and increasing the time step. Here, we improve upon this method to further target GPU accelerated platforms by reducing GPU memory accesses, removing a synchronization point, and better clustering computations. The modification ran over two times faster in some cases even though more computations were required, demonstrating the merit of improving memory handling on the GPU. Furthermore, we discover that the modification also has a near 100% hit rate in fast on-chip L1 cache and discuss the reasons for this. In concluding, we remark on further potential improvements to GPU efficiency.