Benchmarking the Memory Hierarchy of Modern GPUs

Xinxin Mei; Kaiyong Zhao; Chengjian Liu; Xiaowen Chu

doi:10.1007/978-3-662-44917-2_13

Communication Dans Un Congrès Année : 2014

Benchmarking the Memory Hierarchy of Modern GPUs

(1) , (1) , (1) , (1)

Xinxin Mei

Fonction : Auteur
PersonId : 994324

Hong Kong Baptist University

Kaiyong Zhao

Fonction : Auteur
PersonId : 994325

Hong Kong Baptist University

Chengjian Liu

Fonction : Auteur
PersonId : 994326

Hong Kong Baptist University

Xiaowen Chu

Fonction : Auteur
PersonId : 994327

Hong Kong Baptist University

Résumé

Memory access efficiency is a key factor for fully exploiting the computational power of Graphics Processing Units (GPUs). However, many details of the GPU memory hierarchy are not released by the vendors. We propose a novel fine-grained benchmarking approach and apply it on two popular GPUs, namely Fermi and Kepler, to expose the previously unknown characteristics of their memory hierarchies. Specifically, we investigate the structures of different cache systems, such as data cache, texture cache, and the translation lookaside buffer (TLB). We also investigate the impact of bank conflict on shared memory access latency. Our benchmarking results offer a better understanding on the mysterious GPU memory hierarchy, which can help in the software optimization and the modelling of GPU architectures. Our source code and experimental results are publicly available.

Domaines

Informatique [cs]

Fichier principal

978-3-662-44917-2_13_Chapter.pdf (716.95 Ko)

Origine : Fichiers produits par l'(les) auteur(s)

Hal Ifip : Connectez-vous pour contacter le contributeur

https://inria.hal.science/hal-01403075

Soumis le : vendredi 25 novembre 2016-14:28:09

Dernière modification le : jeudi 5 mars 2020-17:40:20

Dates et versions

hal-01403075 , version 1 (25-11-2016)

Licence

Paternité

Identifiants

HAL Id : hal-01403075 , version 1
DOI : 10.1007/978-3-662-44917-2_13

Citer

Xinxin Mei, Kaiyong Zhao, Chengjian Liu, Xiaowen Chu. Benchmarking the Memory Hierarchy of Modern GPUs. 11th IFIP International Conference on Network and Parallel Computing (NPC), Sep 2014, Ilan, Taiwan. pp.144-156, ⟨10.1007/978-3-662-44917-2_13⟩. ⟨hal-01403075⟩

Exporter

BibTeX XML-TEI Dublin Core DC Terms EndNote DataCite

Collections

IFIP-LNCS IFIP IFIP-AICT IFIP-TC IFIP-LNCS-8707 IFIP-TC10 IFIP-NPC IFIP-WG10-3

69 Consultations

158 Téléchargements

Benchmarking the Memory Hierarchy of Modern GPUs

Résumé

Domaines

Dates et versions

Licence

Identifiants

Citer

Exporter

Collections

Altmetric

Partager