Nothing Special   »   [go: up one dir, main page]

skip to main content
10.1109/SC.2005.41acmconferencesArticle/Chapter ViewAbstractPublication PagesscConference Proceedingsconference-collections
Article

Leading Computational Methods on Scalar and Vector HEC Platforms

Published: 12 November 2005 Publication History

Abstract

The last decade has witnessed a rapid proliferation of superscalar cache-based microprocessors to build high-end computing (HEC) platforms, primarily because of their generality, scalability, and cost effectiveness. However, the growing gap between sustained and peak performance for full-scale scientific applications on conventional supercomputers has become a major concern in high performance computing, requiring significantly larger systems and application scalability than implied by peak performance in order to achieve desired performance. The latest generation of custom-built parallel vector systems have the potential to address this issue for numerical algorithms with sufficient regularity in their computational structure. In this work we explore applications drawn from four areas: atmospheric modeling (CAM), magnetic fusion (GTC), plasma physics (LBMHD3D), and material science (PARATEC). We compare performance of the vector-based Cray X1, Earth Simulator, and newly-released NEC SX-8 and Cray X1E, with performance of three leading commodity-based superscalar platforms utilizing the IBM Power3, Intel Itanium2, and AMD Opteron processors. Our work makes several significant contributions: the first reported vector performance results for CAM simulations utilizing a finite-volume dynamical core on a high-resolution atmospheric grid; a new data-decomposition scheme for GTC that (for the first time) enables a breakthrough of the Teraflop barrier; the introduction of a new three-dimensional Lattice Boltzmann magneto-hydrodynamic implementation used to study the onset evolution of plasma turbulence that achieves over 26Tflop/s on 4800 ESpromodity-based superscalar platforms utilizing the IBM Power3, Intel Itanium2, and AMD Opteron processors, with modern parallel vector systems: the Cray X1, Earth Simulator (ES), and the NEC SX-8. Additionally, we examine performance of CAM on the recently-released Cray X1E. Our research team was the first international group to conduct a performance evaluation study at the Earth Simulator Center; remote ES access is not available.Our work builds on our previous efforts [16, 17] and makes several significant contributions: the first reported vector performance results for CAM simulations utilizing a finite-volume dynamical core on a high-resolution atmospheric grid; a new datadecomposition scheme for GTC that (for the first time) enables a breakthrough of the Teraflop barrier; the introduction of a new three-dimensional Lattice Boltzmann magneto-hydrodynamic implementation used to study the onset evolution of plasma turbulence that achieves over 26Tflop/s on 4800 ES processors; and the largest PARATEC cell size atomistic simulation to date. Overall, results show that the vector architectures attain unprecedented aggregate performance across our application suite, demonstrating the tremendous potential of modern parallel vector systems.

References

[1]
{1} A Science-Based Case for Large-Scale Simulation (SCALES). http://www.pnl.gov/scales.
[2]
{2} CAM3.1. http://www.ccsm.ucar.edu/models/atm-cam/.
[3]
{3} HPC challenge benchmark. http://icl.cs.utk.edu/hpcc/index.html.
[4]
{4} ORNL Cray X1 Evaluation. http://www.csm.ornl.gov/~dunigan/cray.
[5]
{5} PARAllel Total Energy Code. http://www.nersc.gov/projects/paratec.
[6]
{6} STREAM: Sustainable memory bandwidth in high performance computer. http://www.cs.virginia.edu/stream.
[7]
{7} W. D. Collins, P. J. Rasch, B. A. Boville, J. J. Hack, J. R. McCaa, D. L. Williamson, B. P. Briegleb, C. M. Bitz, S.-J. Lin, and M. Zhang. The Formulation and Atmospheric Simulation of the Community Atmosphere Model: CAM3. Journal of Climate, to appear, 2005.
[8]
{8} P. J. Dellar. Lattice kinetic schemes for magnetohydrodynamics. J. Comput. Phys., 79, 2002.
[9]
{9} T. H. Dunigan Jr., J. S. Vetter, J. B. White III, and P. H. Worley. Performance evaluation of the Cray X1 distributed shared-memory architecture. IEEE Micro, 25(1):30-40, January/February 2005.
[10]
{10} S. Habata, K. Umezawa, M. Yokokawa, and S. Kitawaki. Hardware system of the Earth Simulator. Parallel Computing, 30:12:1287-1313, 2004.
[11]
{11} W. W. Lee. Gyrokinetic particle simulation model. J. Comp. Phys., 72, 1987.
[12]
{12} S.-J. Lin and R. B. Rood. Multidimensional flux form semi-lagrangian transport schemes. Mon. Wea. Rev., 124: 2046-2070, 1996.
[13]
{13} Z. Lin, T. S. Hahm, W. W. Lee, W. M. Tang, and R. B. White. Turbulent transport reduction by zonal flows: Massively parallel simulations. Science, Sep 1998.
[14]
{14} A. Macnab, G. Vahala, P. Pavlo, L. Vahala, and M. Soe. Lattice boltzmann model for dissipative incompressible MHD. In Proc. 28th EPS Conference on Controlled Fusion and Plasma Physics, volume 25A, 2001.
[15]
{15} A. A. Mirin and W. B. Sawyer. A scalable implemenation of a finite-volume dynamical core in the Community Atmosphere Model. International Journal of High Performance Computing Applications, 19(3), August 2005.
[16]
{16} L. Oliker, A. Canning, J. Carter, J. Shalf, and S. Ethier. Scientific computations on modern parallel vector systems. In Proc. SC2004: High performance computing, networking, and storage conference, 2004.
[17]
{17} L. Oliker, A. Canning, J. Carter, J. Shalf, D. Skinner, S. Ethier, R. Biswas, M. J. Djomehri, and R. F. Van der Wijngaart. Performance evaluation of the SX-6 vector architecture for scientific computations. Concurrency and Computation; Practice and Experience, 17:1:69-93, 2005.
[18]
{18} W. M. Putman, S. J. Lin, and B. Shen. Cross-platform performance of a portable communication module and the NASA finite volume general circulation model. International Journal of High Performance Computing Applications, 19(3), August 2005.
[19]
{19} M. Rancic, R. J. Purser, and F. Mesinger. A global shallow-water model using an expanded spherical cube: gnomic versus conformal coordinates. Q. J. R. Met. Soc., 122: 959-982, 1996.
[20]
{20} D. Skinner. Integrated Performance Monitoring: A portable profiling infrastructure for parallel applications. In Proc. ISC2005: International Supercomputing Conference, volume to appear, Heidelberg, Germany, 2005.
[21]
{21} S. Succi. The lattice boltzmann equation for fluids and beyond. Oxford Science Publ., 2001.
[22]
{22} H. Uehara, M. Tamura, and M. Yokokawa. MPI performance measurement on the Earth Simulator. Technical Report # 15, NEC Research and Development, 2003/1.
[23]
{23} L. W. Wang. Calculating the influence of external charges on the photoluminescence of a CdSe quantum dot. J. Phys. Chem., 105:2360, 2001.
[24]
{24} G. Wellein, T. Zeiser, S. Donath, and G. Hager. On the single processor performance of simple lattice bolzmann kernels. Computers and Fluids, In press.
[25]
{25} P. H. Worley and J. B. Drake. Performance portability in the physical parameterizations of the Community Atmosphere Model. International Journal of High Performance Computing Applications, 19(3):1-15, August 2005.

Cited By

View all
  • (2015)System-Level Support for Composition of ApplicationsProceedings of the 5th International Workshop on Runtime and Operating Systems for Supercomputers10.1145/2768405.2768412(1-8)Online publication date: 16-Jun-2015
  • (2011)Just in timeProceedings of the 20th international symposium on High performance distributed computing10.1145/1996130.1996137(27-36)Online publication date: 8-Jun-2011
  • (2009)Performance evaluation of NEC SX-9 using real science and engineering applicationsProceedings of the Conference on High Performance Computing Networking, Storage and Analysis10.1145/1654059.1654088(1-12)Online publication date: 14-Nov-2009
  • Show More Cited By

Index Terms

  1. Leading Computational Methods on Scalar and Vector HEC Platforms

                      Recommendations

                      Comments

                      Please enable JavaScript to view thecomments powered by Disqus.

                      Information & Contributors

                      Information

                      Published In

                      cover image ACM Conferences
                      SC '05: Proceedings of the 2005 ACM/IEEE conference on Supercomputing
                      November 2005
                      829 pages
                      ISBN:1595930612

                      Sponsors

                      Publisher

                      IEEE Computer Society

                      United States

                      Publication History

                      Published: 12 November 2005

                      Check for updates

                      Qualifiers

                      • Article

                      Conference

                      SC '05
                      Sponsor:

                      Acceptance Rates

                      SC '05 Paper Acceptance Rate 62 of 260 submissions, 24%;
                      Overall Acceptance Rate 1,516 of 6,373 submissions, 24%

                      Upcoming Conference

                      Contributors

                      Other Metrics

                      Bibliometrics & Citations

                      Bibliometrics

                      Article Metrics

                      • Downloads (Last 12 months)1
                      • Downloads (Last 6 weeks)1
                      Reflects downloads up to 21 Nov 2024

                      Other Metrics

                      Citations

                      Cited By

                      View all
                      • (2015)System-Level Support for Composition of ApplicationsProceedings of the 5th International Workshop on Runtime and Operating Systems for Supercomputers10.1145/2768405.2768412(1-8)Online publication date: 16-Jun-2015
                      • (2011)Just in timeProceedings of the 20th international symposium on High performance distributed computing10.1145/1996130.1996137(27-36)Online publication date: 8-Jun-2011
                      • (2009)Performance evaluation of NEC SX-9 using real science and engineering applicationsProceedings of the Conference on High Performance Computing Networking, Storage and Analysis10.1145/1654059.1654088(1-12)Online publication date: 14-Nov-2009
                      • (2009)Event-based systemsProceedings of the Third ACM International Conference on Distributed Event-Based Systems10.1145/1619258.1619261(1-10)Online publication date: 6-Jul-2009
                      • (2009)DataStagerProceedings of the 18th ACM international symposium on High performance distributed computing10.1145/1551609.1551618(39-48)Online publication date: 11-Jun-2009
                      • (2008)Towards Ultra-High Resolution Models of Climate and WeatherInternational Journal of High Performance Computing Applications10.1177/109434200708502322:2(149-165)Online publication date: 1-May-2008
                      • (2008)Performance Analysis of Leading HPC Architectures With Beambeam3DInternational Journal of High Performance Computing Applications10.1177/109434200608502422:1(21-32)Online publication date: 1-Feb-2008
                      • (2008)Scientific Application Performance On Leading Scalar and Vector Supercomputering PlatformsInternational Journal of High Performance Computing Applications10.1177/109434200608502022:1(5-20)Online publication date: 1-Feb-2008
                      • (2007)BXSA for fast processing of scientific dataProceedings of the 2007 spring simulation multiconference - Volume 210.5555/1404680.1404750(441-446)Online publication date: 25-Mar-2007
                      • (2007)The Cray BlackWidowProceedings of the 2007 ACM/IEEE conference on Supercomputing10.1145/1362622.1362646(1-12)Online publication date: 16-Nov-2007

                      View Options

                      Login options

                      View options

                      PDF

                      View or Download as a PDF file.

                      PDF

                      eReader

                      View online with eReader.

                      eReader

                      Media

                      Figures

                      Other

                      Tables

                      Share

                      Share

                      Share this Publication link

                      Share on social media