Benjamin Worpitz scite author profile

Benjamin Worpitz

4Publications

75Citation Statements Received

22Citation Statements Given

How they've been cited

106

How they cite others

Affiliations

LogMeIn (United Kingdom), TU Dresden, Helmholtz-Zentrum Dresden-Rossendorf

Publications

Order By: Most citations

Alpaka -- An Abstraction Library for Parallel Kernel Acceleration

Zenker

Worpitz

Widera

et al. 2016

View full text Add to dashboard Cite

Porting applications to new hardware or programming models is a tedious and error prone process. Every help that eases these burdens is saving developer time that can then be invested into the advancement of the application itself instead of preserving the status-quo on a new platform.The Alpaka library defines and implements an abstract hierarchical redundant parallelism model. The model exploits parallelism and memory hierarchies on a node at all levels available in current hardware. By doing so, it allows to achieve platform and performance portability across various types of accelerators by ignoring specific unsupported levels and utilizing only the ones supported on a specific accelerator. All hardware types (multi-and many-core CPUs, GPUs and other accelerators) are supported for and can be programmed in the same way. The Alpaka C++ template interface allows for straightforward extension of the library to support other accelerators and specialization of its internals for optimization.Running Alpaka applications on a new (and supported) platform requires the change of only one source code line instead of a lot of #ifdefs. * This project has received funding from the European Unions Horizon 2020 research and innovation programme under grant agreement No 654220

show abstract

Tuning and Optimization for a Variety of Many-Core Architectures Without Changing a Single Line of Implementation Code Using the Alpaka Library

Matthes

Widera

Zenker

et al. 2017

View full text Add to dashboard Cite

We present an analysis on optimizing performance of a single C++11 source code using the Alpaka hardware abstraction library. For this we use the general matrix multiplication (GEMM) algorithm in order to show that compilers can optimize Alpaka code effectively when tuning key parameters of the algorithm. We do not intend to rival existing, highly optimized DGEMM versions, but merely choose this example to prove that Alpaka allows for platform-specific tuning with a single source code. In addition we analyze the optimization potential available with vendor-specific compilers when confronted with the heavily templated abstractions of Alpaka. We specifically test the code for bleeding edge architectures such as Nvidia's Tesla P100, Intel's Knights Landing (KNL) and Haswell architecture as well as IBM's Power8 system. On some of these we are able to reach almost 50% of the peak floating point operation performance using the aforementioned means. When adding compilerspecific #pragmas we are able to reach 5 TFLOPs /s on a P100 and over 1 TFLOPs /s on a KNL system.

show abstract

Investigating Performance Portability Of A Highly Scalable Particle-In-Cell Simulation Code On Various Multi-Core Architectures

Worpitz¹,

Juckeland²,

Knüpfer³

et al. 2015

View full text Add to dashboard Cite

PIConGPU 0.5.0: Perfectly Matched Layer (PML) and Bug Fixes

Huebl¹,

Widera²,

Worpitz³

et al. 2020

View full text Add to dashboard Cite

scite is a Brooklyn-based organization that helps researchers better discover and understand research articles through Smart Citations–citations that display the context of the citation and describe whether the article provides supporting or contrasting evidence. scite is used by students and researchers from around the world and is funded in part by the National Science Foundation and the National Institute on Drug Abuse of the National Institutes of Health.

Contact Info

customersupport@researchsolutions.com

10624 S. Eastern Ave., Ste. A-614

Henderson, NV 89052, USA

This site is protected by reCAPTCHA and the Google Privacy Policy and Terms of Service apply.

Blog Terms and Conditions API Terms Privacy Policy Contact Cookie Preferences Do Not Sell or Share My Personal Information

Made with 💙 for researchers

Part of the Research Solutions Family.