Optimization of the Parallel Finite Element Method for the Earth Simulator

N. Kushida, H. Okuda
{"title":"Optimization of the Parallel Finite Element Method for the Earth Simulator","authors":"N. Kushida, H. Okuda","doi":"10.1299/JCST.2.81","DOIUrl":null,"url":null,"abstract":"The feasibility of the GeoFEM as a platform for the parallel finite element method on the earth simulator was investigated. Since the earth simulator consists of 640 SMP nodes, each of which has eight vector processors, there are three levels of hierarchical parallelization methods: inter-node, intra-node, and vectorization. GeoFEM has extremely high inter-node parallel efficiency. However, the application of GeoFEM in an environment involving over 1,000 processors has not yet been examined. Furthermore, the hierarchical architecture of the Earth Simulator requires optimization for intra-node parallelization and vectorization for better practical performance. Various ordering methods have been used to accomplish intra-node parallelization and vectorization, and we eventually achieved a performance of 10 TeraFLOPS for a 6.4-GDOF problem.","PeriodicalId":196913,"journal":{"name":"Journal of Computational Science and Technology","volume":"7 1","pages":"0"},"PeriodicalIF":0.0000,"publicationDate":"1900-01-01","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"6","resultStr":null,"platform":"Semanticscholar","paperid":null,"PeriodicalName":"Journal of Computational Science and Technology","FirstCategoryId":"1085","ListUrlMain":"https://doi.org/10.1299/JCST.2.81","RegionNum":0,"RegionCategory":null,"ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"","JCRName":"","Score":null,"Total":0}
引用次数: 6

Abstract

The feasibility of the GeoFEM as a platform for the parallel finite element method on the earth simulator was investigated. Since the earth simulator consists of 640 SMP nodes, each of which has eight vector processors, there are three levels of hierarchical parallelization methods: inter-node, intra-node, and vectorization. GeoFEM has extremely high inter-node parallel efficiency. However, the application of GeoFEM in an environment involving over 1,000 processors has not yet been examined. Furthermore, the hierarchical architecture of the Earth Simulator requires optimization for intra-node parallelization and vectorization for better practical performance. Various ordering methods have been used to accomplish intra-node parallelization and vectorization, and we eventually achieved a performance of 10 TeraFLOPS for a 6.4-GDOF problem.
查看原文
分享 分享
微信好友 朋友圈 QQ好友 复制链接
本刊更多论文
地球模拟器并行有限元法的优化
研究了GeoFEM在地球模拟器上作为并行有限元方法平台的可行性。由于地球模拟器由640个SMP节点组成,每个节点有8个矢量处理器,因此有三个层次的分层并行化方法:节点间、节点内和矢量化。GeoFEM具有极高的节点间并行效率。但是,GeoFEM在涉及1 000多台处理机的环境中的应用尚未得到审查。此外,地球模拟器的分层结构需要优化节点内并行化和向量化,以获得更好的实际性能。我们使用了各种排序方法来完成节点内并行化和向量化,最终在6.4 gdof问题上实现了10 TeraFLOPS的性能。
本文章由计算机程序翻译,如有差异,请以英文原文为准。
求助全文
约1分钟内获得全文 去求助
来源期刊
自引率
0.00%
发文量
0
期刊最新文献
Design and Optimization of a Gas Burner for TPV Application Experimental and Numerical Approaches for Reliability Evaluation of Electronic Packaging Two-Layer Viscous Shallow-Water Equations and Conservation Laws Lattice Boltzmann Simulation of Two-Phase Viscoelastic Fluid Flows An Inexact Balancing Preconditioner for Large-Scale Structural Analysis
×
引用
GB/T 7714-2015
复制
MLA
复制
APA
复制
导出至
BibTeX EndNote RefMan NoteFirst NoteExpress
×
×
提示
您的信息不完整,为了账户安全,请先补充。
现在去补充
×
提示
您因"违规操作"
具体请查看互助需知
我知道了
×
提示
现在去查看 取消
×
提示
确定
0
微信
客服QQ
Book学术公众号 扫码关注我们
反馈
×
意见反馈
请填写您的意见或建议
请填写您的手机或邮箱
已复制链接
已复制链接
快去分享给好友吧!
我知道了
×
扫码分享
扫码分享
Book学术官方微信
Book学术文献互助
Book学术文献互助群
群 号:481959085
Book学术
文献互助 智能选刊 最新文献 互助须知 联系我们:info@booksci.cn
Book学术提供免费学术资源搜索服务,方便国内外学者检索中英文文献。致力于提供最便捷和优质的服务体验。
Copyright © 2023 Book学术 All rights reserved.
ghs 京公网安备 11010802042870号 京ICP备2023020795号-1