计算机科学
数据流
多线程
数据流体系结构
计算机体系结构
现场可编程门阵列
云计算
并行计算
线程(计算)
覆盖
建筑
延迟(音频)
运行时系统
嵌入式系统
操作系统
艺术
视觉艺术
电信
作者
Lucas Bragança Da Silva,Ricardo Ferreira,Michael Canesche,Marcelo M. Menezes,Maria de Fátima Araújo Vieira,Jeronimo Costa Penha,Peter Jamieson,José Augusto M. Nacif
出处
期刊:ACM Transactions in Embedded Computing Systems
[Association for Computing Machinery]
日期:2019-10-07
卷期号:18 (5s): 1-20
被引量:18
摘要
In this work, we propose a framework called REconfigurable Accelerator DeploY (READY), the first framework to support polynomial runtime mapping of dataflow applications in high-performance CPU-FPGA platforms. READY introduces an efficient mapping with fine-grained multithreading onto an overlay architecture that hides the latency of a global interconnection network. In addition to our overlay architecture, we show how this system helps solve some of the challenges for FPGA cloud computing adoption in high-performance computing. The framework encapsulates dataflow descriptions by using a target independent, high-level API, and a dataflow model that allows for explicit spatial and temporal parallelism. READY directly maps the dataflow kernels onto the accelerator. Our tool is flexible and extensible and provides the infrastructure to explore different accelerator designs. We validate READY on the Intel Harp platform, and our experimental results show an average 2x execution runtime improvement when compared to an 8-thread multi-core processor.
科研通智能强力驱动
Strongly Powered by AbleSci AI