kth.sePublications KTH
Change search
CiteExportLink to record
Permanent link

Direct link
Cite
Citation style
  • apa
  • ieee
  • modern-language-association-8th-edition
  • vancouver
  • Other style
More styles
Language
  • de-DE
  • en-GB
  • en-US
  • fi-FI
  • nn-NO
  • nn-NB
  • sv-SE
  • Other locale
More languages
Output format
  • html
  • text
  • asciidoc
  • rtf
Multi-Partner Project: Multi-GPU Performance Portability Analysis for CFD Simulations at Scale
National Technical University of Athens, Greece.
National Technical University of Athens, Greece.
National Technical University of Athens, Greece.
KTH, School of Engineering Sciences (SCI), Engineering Mechanics, Fluid Mechanics.ORCID iD: 0000-0003-0769-9101
Show others and affiliations
2026 (English)In: 2026 Design, Automation and Test in Europe Conference, DATE 2026 - Proceedings, Institute of Electrical and Electronics Engineers (IEEE) , 2026Conference paper, Published paper (Refereed)
Abstract [en]

As heterogeneous supercomputing architectures leveraging GPUs become increasingly central to high-performance computing (HPC), it is crucial for computational fluid dynamics (CFD) simulations, a de-facto HPC workload, to efficiently utilize such hardware. One of the key challenges of HPC codes is performance portability, i.e. the ability to maintain near-optimal performance across different accelerators. In the context of the REFMAP project, which targets scalable, GPU-enabled multi-fidelity CFD for urban airflow prediction, this paper analyzes the performance portability of SOD2D, a state-of-the-art Spectral Elements simulation framework across AMD and NVIDIA GPU architectures. We first discuss the physical and numerical models underlying SOD2D, highlighting its computational hotspots. Then, we examine its performance and scalability in a multi-level manner, i.e. defining and characterizing an extensive full-stack design space spanning across application, software and hardware infrastructure related parameters. Single-GPU performance characterization across server-grade NVIDIA and AMD GPU architectures and vendor-specific compiler stacks, show the potential as well as the diverse effect of memory access optimizations, i.e. 0.69× - 3.91× deviations in acceleration speedup. Performance variability of SOD2D at scale is further examined on the LUMI multi-GPU cluster, where profiling reveals similar throughput variations, highlighting the limits of performance projections and the need for multi-level, informed tuning.

Place, publisher, year, edition, pages
Institute of Electrical and Electronics Engineers (IEEE) , 2026.
Keywords [en]
CFD, Performance portability, Spectral Finite Element Method (FEM), design space exploration, high-fidelity simulation, multi-GPU acceleration, scalability analysis
National Category
Computer Sciences Computer Systems
Identifiers
URN: urn:nbn:se:kth:diva-384140DOI: 10.23919/DATE69613.2026.11539345Scopus ID: 2-s2.0-105041993384OAI: oai:DiVA.org:kth-384140DiVA, id: diva2:2079859
Conference
2026 Design, Automation and Test in Europe Conference, DATE 2026, Verona, Italy, April 20-22, 2026
Note

Part of ISBN 9783982674117

QC 20260625

Available from: 2026-06-25 Created: 2026-06-25 Last updated: 2026-06-25Bibliographically approved

Open Access in DiVA

No full text in DiVA

Other links

Publisher's full textScopus

Authority records

Umair, MohammadVincent, JonathanGong, JingZampino, Gerardo

Search in DiVA

By author/editor
Umair, MohammadVincent, JonathanGong, JingZampino, Gerardo
By organisation
Fluid MechanicsCentre for High Performance Computing, PDC
Computer SciencesComputer Systems

Search outside of DiVA

GoogleGoogle Scholar

doi
urn-nbn

Altmetric score

doi
urn-nbn
Total: 7 hits
CiteExportLink to record
Permanent link

Direct link
Cite
Citation style
  • apa
  • ieee
  • modern-language-association-8th-edition
  • vancouver
  • Other style
More styles
Language
  • de-DE
  • en-GB
  • en-US
  • fi-FI
  • nn-NO
  • nn-NB
  • sv-SE
  • Other locale
More languages
Output format
  • html
  • text
  • asciidoc
  • rtf