CMD: An Integrated CGRA Framework with Cluster-Based Distributed Memory Design

📅 2026-09-05
📈 Citations: 0
Influential: 0
📄 PDF
📝 Abstract
Coarse-Grained Reconfigurable Arrays (CGRAs) are a promising solution for achieving high energy efficiency and reconfigurability across various application domains, but their performance is often crippled by rigid memory architectures that limit the number and location of tiles that can access data memory. This creates a significant bottleneck for kernels with intensive memory accesses. To address this, we propose CMD, an integrated CGRA framework featuring cluster-based distributed memory design with a co-designed compilation toolchain. The compiler includes a novel memory-aware mapper and a design space exploration (DSE) mechanism that identifies the optimal memory architecture design for specific kernels. Experimental results show that our post-DSE CMD CGRAs achieve an average speedup of $1.39\times$ over a conventional CGRA while simultaneously reducing the total area to an average of $0.912\times$ of the conventional CGRA.
Problem

Research questions and friction points this paper is trying to address.

CGRAs
memory architecture
performance bottleneck
Innovation

Methods, ideas, or system contributions that make the work stand out.

Cluster-Based Distributed Memory
Memory-Aware Mapper
Design Space Exploration (DSE)
💼 Related Jobs
No related jobs found.
S
Shangkun Li
The Hong Kong University of Science and Technology, Hong Kong SAR, China
C
Cheng Tan
Google, Mountain View, USA
Zeyu Li
Zeyu Li
Hong Kong University of Science and Technology(Guang Zhou)
GPUHigh Performance Compute
J
Jinming Ge
The Hong Kong University of Science and Technology, Hong Kong SAR, China
J
Jiawei Liang
The Hong Kong University of Science and Technology, Hong Kong SAR, China
H
Hao Yang
The George Washington University, Washington, D.C., USA
L
Linfeng Du
The Hong Kong University of Science and Technology, Hong Kong SAR, China
Jiang Xu
Jiang Xu
Hong Kong University of Science and Technology (Guangzhou)
MPSoCNetwork on ChipHW/SW-CodesignOptical Neural Network
Wei Zhang
Wei Zhang
Hong Kong University of Science and Technology
Embedded systemReconfigurable computingMulticore systemNanoelectronics