An automatic input-sensitive approach for heterogeneous task partitioning
dc.contributor.author | Kofler, Kofler | |
dc.contributor.author | Grasso, Ivan | |
dc.contributor.author | Cosenza, Biagio | |
dc.contributor.author | Fahringer, Thomas | |
dc.date.accessioned | 2017-10-26T10:38:33Z | |
dc.date.available | 2017-10-26T10:38:33Z | |
dc.date.issued | 2013 | |
dc.description.abstract | Unleashing the full potential of heterogeneous systems, consisting of multi-core CPUs and GPUs, is a challenging task due to the difference in processing capabilities, memory availability, and communication latencies of different computational resources. In this paper we propose a novel approach that automatically optimizes task partitioning for different (input) problem sizes and different heterogeneous multi-core architectures. We use the Insieme source-to-source compiler to translate a single-device OpenCL program into a multi-device OpenCL program. The Insieme Runtime System then performs dynamic task partitioning based on an offline-generated prediction model. In order to derive the prediction model, we use a machine learning approach based on Artificial Neural Networks (ANN) that incorporates static program features as well as dynamic, input sensitive features. Principal component analysis have been used to further improve the task partitioning. Our approach has been evaluated over a suite of 23 programs and respectively achieves a performance improvement of 22% and 25% compared to an execution of the benchmarks on a single CPU and a single GPU which is equal to 87.5% of the optimal performance. | en |
dc.identifier.isbn | 978-1-4503-2130-3 | |
dc.identifier.uri | https://depositonce.tu-berlin.de/handle/11303/7020 | |
dc.identifier.uri | http://dx.doi.org/10.14279/depositonce-6341 | |
dc.language.iso | en | |
dc.rights.uri | http://rightsstatements.org/vocab/InC/1.0/ | |
dc.subject.ddc | 004 Datenverarbeitung; Informatik | |
dc.subject.other | code analysis | en |
dc.subject.other | compilers | en |
dc.subject.other | gpu | en |
dc.subject.other | heterogeneous computing | en |
dc.subject.other | machine learning | en |
dc.subject.other | runtime system | en |
dc.subject.other | task partitioning | en |
dc.title | An automatic input-sensitive approach for heterogeneous task partitioning | en |
dc.type | Conference Object | en |
dc.type.version | acceptedVersion | en |
dcterms.bibliographicCitation.doi | 10.1145/2464996.2465007 | |
dcterms.bibliographicCitation.originalpublishername | Association for Computing Machinery (ACM) | en |
dcterms.bibliographicCitation.originalpublisherplace | New York, NY | en |
dcterms.bibliographicCitation.pageend | 160 | |
dcterms.bibliographicCitation.pagestart | 149 | |
dcterms.bibliographicCitation.proceedingstitle | Proceedings of the 27th International ACM Conference on International Conference on Supercomputing | en |
tub.accessrights.dnb | domain | |
tub.affiliation | Fak. 4 Elektrotechnik und Informatik::Inst. Technische Informatik und Mikroelektronik::FG Architektur eingebetteter Systeme | de |
tub.affiliation.faculty | Fak. 4 Elektrotechnik und Informatik | de |
tub.affiliation.group | FG Architektur eingebetteter Systeme | de |
tub.affiliation.institute | Inst. Technische Informatik und Mikroelektronik | de |
tub.publisher.universityorinstitution | Technische Universität Berlin | en |
Files
Original bundle
1 - 1 of 1