地球模拟器并行有限元法的优化

N. Kushida, H. Okuda
{"title":"地球模拟器并行有限元法的优化","authors":"N. Kushida, H. Okuda","doi":"10.1299/JCST.2.81","DOIUrl":null,"url":null,"abstract":"The feasibility of the GeoFEM as a platform for the parallel finite element method on the earth simulator was investigated. Since the earth simulator consists of 640 SMP nodes, each of which has eight vector processors, there are three levels of hierarchical parallelization methods: inter-node, intra-node, and vectorization. GeoFEM has extremely high inter-node parallel efficiency. However, the application of GeoFEM in an environment involving over 1,000 processors has not yet been examined. Furthermore, the hierarchical architecture of the Earth Simulator requires optimization for intra-node parallelization and vectorization for better practical performance. Various ordering methods have been used to accomplish intra-node parallelization and vectorization, and we eventually achieved a performance of 10 TeraFLOPS for a 6.4-GDOF problem.","PeriodicalId":196913,"journal":{"name":"Journal of Computational Science and Technology","volume":"7 1","pages":"0"},"PeriodicalIF":0.0000,"publicationDate":"1900-01-01","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"6","resultStr":"{\"title\":\"Optimization of the Parallel Finite Element Method for the Earth Simulator\",\"authors\":\"N. Kushida, H. Okuda\",\"doi\":\"10.1299/JCST.2.81\",\"DOIUrl\":null,\"url\":null,\"abstract\":\"The feasibility of the GeoFEM as a platform for the parallel finite element method on the earth simulator was investigated. Since the earth simulator consists of 640 SMP nodes, each of which has eight vector processors, there are three levels of hierarchical parallelization methods: inter-node, intra-node, and vectorization. GeoFEM has extremely high inter-node parallel efficiency. However, the application of GeoFEM in an environment involving over 1,000 processors has not yet been examined. Furthermore, the hierarchical architecture of the Earth Simulator requires optimization for intra-node parallelization and vectorization for better practical performance. Various ordering methods have been used to accomplish intra-node parallelization and vectorization, and we eventually achieved a performance of 10 TeraFLOPS for a 6.4-GDOF problem.\",\"PeriodicalId\":196913,\"journal\":{\"name\":\"Journal of Computational Science and Technology\",\"volume\":\"7 1\",\"pages\":\"0\"},\"PeriodicalIF\":0.0000,\"publicationDate\":\"1900-01-01\",\"publicationTypes\":\"Journal Article\",\"fieldsOfStudy\":null,\"isOpenAccess\":false,\"openAccessPdf\":\"\",\"citationCount\":\"6\",\"resultStr\":null,\"platform\":\"Semanticscholar\",\"paperid\":null,\"PeriodicalName\":\"Journal of Computational Science and Technology\",\"FirstCategoryId\":\"1085\",\"ListUrlMain\":\"https://doi.org/10.1299/JCST.2.81\",\"RegionNum\":0,\"RegionCategory\":null,\"ArticlePicture\":[],\"TitleCN\":null,\"AbstractTextCN\":null,\"PMCID\":null,\"EPubDate\":\"\",\"PubModel\":\"\",\"JCR\":\"\",\"JCRName\":\"\",\"Score\":null,\"Total\":0}","platform":"Semanticscholar","paperid":null,"PeriodicalName":"Journal of Computational Science and Technology","FirstCategoryId":"1085","ListUrlMain":"https://doi.org/10.1299/JCST.2.81","RegionNum":0,"RegionCategory":null,"ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"","JCRName":"","Score":null,"Total":0}
引用次数: 6

摘要

研究了GeoFEM在地球模拟器上作为并行有限元方法平台的可行性。由于地球模拟器由640个SMP节点组成,每个节点有8个矢量处理器,因此有三个层次的分层并行化方法:节点间、节点内和矢量化。GeoFEM具有极高的节点间并行效率。但是,GeoFEM在涉及1 000多台处理机的环境中的应用尚未得到审查。此外,地球模拟器的分层结构需要优化节点内并行化和向量化,以获得更好的实际性能。我们使用了各种排序方法来完成节点内并行化和向量化,最终在6.4 gdof问题上实现了10 TeraFLOPS的性能。
本文章由计算机程序翻译,如有差异,请以英文原文为准。
Optimization of the Parallel Finite Element Method for the Earth Simulator
The feasibility of the GeoFEM as a platform for the parallel finite element method on the earth simulator was investigated. Since the earth simulator consists of 640 SMP nodes, each of which has eight vector processors, there are three levels of hierarchical parallelization methods: inter-node, intra-node, and vectorization. GeoFEM has extremely high inter-node parallel efficiency. However, the application of GeoFEM in an environment involving over 1,000 processors has not yet been examined. Furthermore, the hierarchical architecture of the Earth Simulator requires optimization for intra-node parallelization and vectorization for better practical performance. Various ordering methods have been used to accomplish intra-node parallelization and vectorization, and we eventually achieved a performance of 10 TeraFLOPS for a 6.4-GDOF problem.
求助全文
通过发布文献求助,成功后即可免费获取论文全文。 去求助
来源期刊
自引率
0.00%
发文量
0
×
引用
GB/T 7714-2015
复制
MLA
复制
APA
复制
导出至
BibTeX EndNote RefMan NoteFirst NoteExpress
×
提示
您的信息不完整,为了账户安全,请先补充。
现在去补充
×
提示
您因"违规操作"
具体请查看互助需知
我知道了
×
提示
确定
请完成安全验证×
copy
已复制链接
快去分享给好友吧!
我知道了
右上角分享
点击右上角分享
0
联系我们:info@booksci.cn Book学术提供免费学术资源搜索服务,方便国内外学者检索中英文文献。致力于提供最便捷和优质的服务体验。 Copyright © 2023 布克学术 All rights reserved.
京ICP备2023020795号-1
ghs 京公网安备 11010802042870号
Book学术文献互助
Book学术文献互助群
群 号:481959085
Book学术官方微信