A Memory Management System Optimized for BDMPI's Memory and Execution Model

Proceedings of the 22nd European MPI Users' Group Meeting Pub Date : 2015-09-21 DOI:10.1145/2802658.2802666

J. Iverson, G. Karypis

{"title":"A Memory Management System Optimized for BDMPI's Memory and Execution Model","authors":"J. Iverson, G. Karypis","doi":"10.1145/2802658.2802666","DOIUrl":null,"url":null,"abstract":"There is a growing need to perform large computations on small systems, as access to large systems is not widely available and cannot keep up with the scaling of data. BDMPI was recently introduced as a way of achieving this for applications written in MPI. BDMPI allows the efficient execution of standard MPI programs on systems whose aggregate amount of memory is smaller than that required by the computations and significantly outperforms other approaches. In this paper we present a virtual memory subsystem which we implemented as part of the BDMPI runtime. Our new virtual memory subsystem, which we call SBMA, bypasses the operating system virtual memory manager to take advantage of BDMPI's node-level cooperative multi-taking. Benchmarking using a synthetic application shows that for the use cases relevant to BDMPI, the overhead incurred by the BDMPI-SBMA system is amortized such that it performs as fast as explicit data movement by the application developer. Furthermore, we tested SBMA with three different classes of applications and our results show that with no modification to the original MPI program, speedups from 2×--12× over a standard BDMPI implementation can be achieved for the included applications.","PeriodicalId":365272,"journal":{"name":"Proceedings of the 22nd European MPI Users' Group Meeting","volume":"50 1","pages":"0"},"PeriodicalIF":0.0000,"publicationDate":"2015-09-21","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"0","resultStr":null,"platform":"Semanticscholar","paperid":null,"PeriodicalName":"Proceedings of the 22nd European MPI Users' Group Meeting","FirstCategoryId":"1085","ListUrlMain":"https://doi.org/10.1145/2802658.2802666","RegionNum":0,"RegionCategory":null,"ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"","JCRName":"","Score":null,"Total":0}

引用次数: 0

Abstract

There is a growing need to perform large computations on small systems, as access to large systems is not widely available and cannot keep up with the scaling of data. BDMPI was recently introduced as a way of achieving this for applications written in MPI. BDMPI allows the efficient execution of standard MPI programs on systems whose aggregate amount of memory is smaller than that required by the computations and significantly outperforms other approaches. In this paper we present a virtual memory subsystem which we implemented as part of the BDMPI runtime. Our new virtual memory subsystem, which we call SBMA, bypasses the operating system virtual memory manager to take advantage of BDMPI's node-level cooperative multi-taking. Benchmarking using a synthetic application shows that for the use cases relevant to BDMPI, the overhead incurred by the BDMPI-SBMA system is amortized such that it performs as fast as explicit data movement by the application developer. Furthermore, we tested SBMA with three different classes of applications and our results show that with no modification to the original MPI program, speedups from 2×--12× over a standard BDMPI implementation can be achieved for the included applications.

查看原文

微信好友朋友圈 QQ好友复制链接

本刊更多论文

一个针对BDMPI的内存和执行模型进行优化的内存管理系统

在小型系统上执行大型计算的需求越来越大，因为对大型系统的访问并不广泛，并且无法跟上数据的扩展。BDMPI最近被引入，作为用MPI编写的应用程序实现这一目标的一种方式。BDMPI允许在内存总量小于计算所需的系统上有效地执行标准MPI程序，并且显著优于其他方法。在本文中，我们提出了一个虚拟内存子系统，作为BDMPI运行时的一部分实现。我们的新虚拟内存子系统(我们称之为SBMA)绕过操作系统虚拟内存管理器来利用BDMPI的节点级协作多占用。使用合成应用程序进行基准测试表明，对于与BDMPI相关的用例，BDMPI- sbma系统产生的开销被分摊，从而使其执行速度与应用程序开发人员的显式数据移动一样快。此外，我们用三种不同类型的应用程序测试了SBMA，结果表明，在不修改原始MPI程序的情况下，对于所包含的应用程序，可以实现比标准BDMPI实现的速度提高2倍至12倍。

本文章由计算机程序翻译，如有差异，请以英文原文为准。

求助全文

约1分钟内获得全文去求助

来源期刊

Proceedings of the 22nd European MPI Users' Group Meeting

自引率

0.00%

发文量

期刊最新文献

Detecting Silent Data Corruption for Extreme-Scale MPI Applications Correctness Analysis of MPI-3 Non-Blocking Communications in PARCOACH Sliding Substitution of Failed Nodes MPI-focused Tracing with OTFX: An MPI-aware In-memory Event Tracing Extension to the Open Trace Format 2 STCI: Scalable RunTime Component Infrastructure