Abstract
We describe the implementation of a thermal compressible Lattice Boltzmann algorithm on an NVIDIA Tesla C2050 system based on the Fermi GP-GPU. We consider two different versions, including and not including reactive effects. We describe the overall organization of the algorithm and give details on its implementations. Efficiency ranges from 25% to 31% of the double precision peak performance of the GP-GPU. We compare our results with a different implementation of the same algorithm, developed and optimized for many-core Intel Westmere CPUs.
| Original language | English |
|---|---|
| Pages (from-to) | 55-62 |
| Journal | Computers & Fluids |
| Volume | 80 |
| DOIs | |
| Publication status | Published - 2013 |
Fingerprint
Dive into the research topics of 'An optimized D2Q37 Lattice Boltzmann code on GP-GPUs'. Together they form a unique fingerprint.Cite this
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver