GpuTejas: A parallel simulator for GPU architectures

2014 21st International Conference on High Performance Computing (HiPC) Pub Date : 2014-12-01 DOI:10.1109/HiPC.2014.7116897

Geetika Malhotra, Seep Goel, S. Sarangi

引用次数: 22

Abstract

In this paper, we introduce a new Java-based parallel GPGPU simulator, GpuTejas. GpuTejas is a fast trace driven simulator, which uses relaxed synchronization, and non-blocking data structures to derive its speedups. Secondly, it introduces a novel scheduling and partitioning scheme for parallelizing a GPU simulator. We evaluate the performance of our simulator with a set of Rodinia benchmarks. We demonstrate a mean speedup of 17.33x with 64 threads over sequential execution, and a speedup of 429X over the widely used simulator GPGPU-Sim. We validated our timing and simulation model by comparing our results with a native system (NVIDIA Tesla M2070). As compared to the sequential version of GpuTejas, the parallel version has an error limited to <;7.67% for our suite of benchmarks, which is similar to the numbers reported by competing parallel simulators.

查看原文

微信好友朋友圈 QQ好友复制链接

本刊更多论文

GpuTejas: GPU架构的并行模拟器

本文介绍了一种新的基于java的并行GPGPU模拟器GpuTejas。GpuTejas是一种快速跟踪驱动模拟器，它使用宽松的同步和非阻塞数据结构来获得其速度。其次，介绍了一种新的GPU模拟器并行化调度和分区方案。我们用一组Rodinia基准来评估模拟器的性能。我们演示了64线程顺序执行的平均加速速度为17.33x，在广泛使用的模拟器GPGPU-Sim上的加速速度为429X。我们通过将我们的结果与本地系统(NVIDIA Tesla M2070)进行比较来验证我们的时序和仿真模型。与连续版本的GpuTejas相比，在我们的基准测试套件中，并行版本的误差限制在< 7.67%，这与竞争的并行模拟器报告的数字相似。

本文章由计算机程序翻译，如有差异，请以英文原文为准。

求助全文

约1分钟内获得全文去求助

来源期刊

2014 21st International Conference on High Performance Computing (HiPC)

自引率

0.00%

发文量