PIMS: A Lightweight Processing-In-Memory Accelerator for Stencil Computations

Name: PIMS: A Lightweight Processing-In-Memory Accelerator for Stencil Computations
Start: 2019-10-01T00:00:00Z
End: 2019-10-03T00:00:00Z
Location: Washington DC, USA

Jie Li, Xi Wang, Antonino Tumeo, Brody Williams, John D. Leidel, Yong Chen

MemSys2019

Abstract

Stencil computation is a classic computational kernel present in many high-performance scientific applications, like image processing and partial differential equation solvers (PDE). A stencil computation sweeps over a multi-dimensional grid and repeatedly updates values associated with points using the values from neighboring points. Stencil computations often employ large datasets that exceed cache capacity, leading to excessive accesses to the memory subsystem. As such, 3D stencil computations on large grid sizes are memory-bound. In this paper we present PIMS, an in-memory accelerator for stencil computations. PIMS, implemented in the logic layer of a 3D stacked memory, exploits the high bandwidth provided by through silicon vias to reduce redundant memory traffic. Our comprehensive evaluation using three different grid sizes with six categories of orders indicate that the proposed architecture reduces 48.25% of data movement on average and obtains up to 65.55% of bank conflict reduction.

Date

Oct 1, 2019 12:00 AM — Oct 3, 2019 12:00 AM

Event

MEMSYS 2019

Location

Washington DC, USA

Stencil Processing-in-Memory 3D-Stacked Memory Computer Architecture HPC

PIMS: A Lightweight Processing-In-Memory Accelerator for Stencil Computations

Abstract

Jie Li

Ph.D. candidate in Computer Science