The FLAME approach: From dense linear algebra algorithms to high-performance multi-accelerator implementations

کد مقاله	کد نشریه	سال انتشار	مقاله انگلیسی	نسخه تمام متن
431537	688570	2012	10 صفحه PDF	دانلود رایگان

عنوان انگلیسی مقاله ISI

دانلود مقاله + سفارش ترجمه

دانلود مقاله ISI انگلیسی

رایگان برای ایرانیان

کلمات کلیدی

Runtime systems - سیستم های زمان اجرا High performance computing - محاسبات با کارایی بالا Graphics processors - پردازنده های گرافیکی

موضوعات مرتبط

مهندسی و علوم پایه مهندسی کامپیوتر نظریه محاسباتی و ریاضیات

پیش نمایش صفحه اول مقاله

The FLAME approach: From dense linear algebra algorithms to high-performance multi-accelerator implementations

چکیده انگلیسی

Parallel accelerators are playing an increasingly important role in scientific computing. However, it is perceived that their weakness nowadays is their reduced “programmability” in comparison with traditional general-purpose CPUs. For the domain of dense linear algebra, we demonstrate that this is not necessarily the case. We show how the libflame library carefully layers routines and abstracts details related to storage and computation, so that extending it to take advantage of multiple accelerators is achievable without introducing platform specific complexity into the library code base. We focus on the experience of the library developer as he develops a library routine for a new operation, reduction of a generalized Hermitian positive definite eigenvalue problem to a standard Hermitian form, and configures the library to target a multi-GPU platform. It becomes obvious that the library developer does not need to know about the parallelization or the details of the multi-accelerator platform. Excellent performance on a system with four NVIDIA Tesla C2050 GPUs is reported. This makes libflame the first library to be released that incorporates multi-GPU functionality for dense matrix computations, setting a new standard for performance.

► Runtime systems for automatic parallelization of dense linear algebra routines.
► Runtime systems for multi-GPU architectures.
► Algorithm-by-blocks for reducing a generalized eigenvalue problem to a standard form.

ناشر

Database: Elsevier - ScienceDirect (ساینس دایرکت)
Journal: Journal of Parallel and Distributed Computing - Volume 72, Issue 9, September 2012, Pages 1134–1143

نویسندگان

Francisco D. Igual, Ernie Chan, Enrique S. Quintana-Ortí, Gregorio Quintana-Ortí, Robert A. van de Geijn, Field G. Van Zee,

علوم انسانی و هنر

فنی، مهندسی و علوم پایه

پزشکی و سلامت

بیو تکنولوژی

پذیرش سفارش ترجمه

The FLAME approach: From dense linear algebra algorithms to high-performance multi-accelerator implementations

دسترسی سریع

ارتباط

English Website