Abstract
In this paper we evaluate two life science algorithms, namely Needleman-Wunsch sequence alignment and Direct Coulomb Summation, for GPUs. Whereas for Needleman-Wunsch it is difficult to get good performance numbers, Direct Coulomb Summation is particularly suitable for graphics cards. We present several optimization techniques, analyze the theoretical potential of the optimizations with respect to the algorithms, and measure the effect on execution times. We target the recent NVIDIA Fermi architecture to evaluate the performance impacts of novel hardware features like the cache subsystem on optimizing transformations. We compare the execution times of CUDA and OpenCL code versions for Fermi and predecessor models with parallel OpenMP versions executed on the main CPU.
| Originalsprache | Englisch |
|---|---|
| Titel | Parallel, Distributed and Network-Based Processing (PDP), 2012 20th Euromicro International Conference on Computing & Processing (Hardware/Software) |
| Verlag | IEEE Computer Society |
| Seiten | 376-383 |
| Seitenumfang | 8 |
| Publikationsstatus | Veröffentlicht - 2012 |
| Veranstaltung | 20th Euromicro International Conference on Parallel, Distributed and Network-based Processing - , Deutschland Dauer: 15 Feb. 2012 → 17 Feb. 2012 |
Konferenz
| Konferenz | 20th Euromicro International Conference on Parallel, Distributed and Network-based Processing |
|---|---|
| Land/Gebiet | Deutschland |
| Zeitraum | 15/02/12 → 17/02/12 |
ÖFOS 2012
- 102022 Softwareentwicklung
- 202031 Netzwerktechnik
Fingerprint
Untersuchen Sie die Forschungsthemen von „Optimization Techniques and Performance Analyses of two Life Science Algorithms for Novel GPU Architectures“. Zusammen bilden sie einen einzigartigen Fingerprint.Zitationsweisen
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver