kth.sePublikationer KTH
Ändra sökning
RefereraExporteraLänk till posten
Permanent länk

Direktlänk
Referera
Referensformat
  • apa
  • ieee
  • modern-language-association-8th-edition
  • vancouver
  • Annat format
Fler format
Språk
  • de-DE
  • en-GB
  • en-US
  • fi-FI
  • nn-NO
  • nn-NB
  • sv-SE
  • Annat språk
Fler språk
Utmatningsformat
  • html
  • text
  • asciidoc
  • rtf
Less for more: reducing intra-cgra connectivity for higher performance and efficiency in hpc
Center for Computational Science (R-CCS), RIKEN, Japan.
Center for Computational Science (R-CCS), RIKEN, Japan.
Center for Computational Science (R-CCS), RIKEN, Japan.
Center for Computational Science (R-CCS), RIKEN, Japan.
Visa övriga samt affilieringar
2023 (Engelska)Ingår i: 2023 IEEE International Parallel and Distributed Processing Symposium Workshops, IPDPSW 2023, Institute of Electrical and Electronics Engineers (IEEE) , 2023, s. 452-459Konferensbidrag, Publicerat paper (Refereegranskat)
Abstract [en]

Coarse-Grained Reconfigurable Arrays (CGRAs) are a class of reconfigurable architectures that inherit the performance of Domain-specific accelerators and the reconfigurability aspects of Field-Programmable Gate Arrays (FPGAs). Historically, CGRAs have been successfully used to accelerate embedded applications and are now considered to accelerate High-Performance Computing (HPC) applications in future supercomputers. However, embedded systems and supercomputers are two vastly different domains with different applications and constraints, and it is today not fully understood what CGRA design decisions adequately cater to the HPC market. One such unknown design decision is regarding the interconnect that facilitates intra-CGRA communication. Our findings show that even the typical king-style mesh-like topology is often under-utilized with a typical HPC workload, leading to inefficiency. This research aims to explore the provisioning of intra-CGRA interconnect for HPC-oriented workloads and, ultimately, recoup the potential performance and efficiency lost by reducing the interconnect complexity. We proposed several reduced interconnect topologies based on the usage statistic. Then we evaluate the tradeoffs regarding hardware cost, routability of DFGs, and computational throughput.

Ort, förlag, år, upplaga, sidor
Institute of Electrical and Electronics Engineers (IEEE) , 2023. s. 452-459
Nyckelord [en]
CGRA, Design space exploration, HPC, Routing architecture, RTL simulation
Nationell ämneskategori
Inbäddad systemteknik
Identifikatorer
URN: urn:nbn:se:kth:diva-336739DOI: 10.1109/IPDPSW59300.2023.00077ISI: 001055030700056Scopus ID: 2-s2.0-85169299919OAI: oai:DiVA.org:kth-336739DiVA, id: diva2:1798497
Konferens
2023 IEEE International Parallel and Distributed Processing Symposium Workshops, IPDPSW 2023, St. Petersburg, United States of America, May 15 2023 - May 19 2023
Anmärkning

Part of ISBN 9798350311990 

QC 20230919

Tillgänglig från: 2023-09-19 Skapad: 2023-09-19 Senast uppdaterad: 2023-10-02Bibliografiskt granskad

Open Access i DiVA

Fulltext saknas i DiVA

Övriga länkar

Förlagets fulltextScopus

Person

Podobas, Artur

Sök vidare i DiVA

Av författaren/redaktören
Podobas, Artur
Av organisationen
Beräkningsvetenskap och beräkningsteknik (CST)
Inbäddad systemteknik

Sök vidare utanför DiVA

GoogleGoogle Scholar

doi
urn-nbn

Altmetricpoäng

doi
urn-nbn
Totalt: 371 träffar
RefereraExporteraLänk till posten
Permanent länk

Direktlänk
Referera
Referensformat
  • apa
  • ieee
  • modern-language-association-8th-edition
  • vancouver
  • Annat format
Fler format
Språk
  • de-DE
  • en-GB
  • en-US
  • fi-FI
  • nn-NO
  • nn-NB
  • sv-SE
  • Annat språk
Fler språk
Utmatningsformat
  • html
  • text
  • asciidoc
  • rtf