An efficient FPGA overlay for MPI-2 RMA parallel applications
Mathieu Leonel, Roland Christian Gamom Ngounou Ewo, Julien Denoulet, Paulin Melatagia Yonta, Bertrand Granado
Centre National de la Recherche Scientifique Université de Yaoundé I Sorbonne Université Unité de Modélisation Mathématique et Informatique des Systèmes Complexes
内容与影响
Design productivity issues, including difficult hardware design and long compile times, are major barriers to the widespread adoption of FPGA-based accelerations in main-stream computing. Enabling virtualized execution of software and hardware tasks on FPGA platforms make them more accessible would to application developers accustomed to software API abstractions such as MPI and fast development cycles. In this work, we show that the MATIP platform provides a viable and efficient FPGA overlay architecture for the design of MPI parallel applications. We support this with a parallel model implementation of a feature extraction algorithm for tone language recognition, which is shown to be at least 7 times more efficient than a C++ MPI-2 RMA implementation of the same parallel model on a CPU and almost 3 times more efficient than a naive FPGA IP implementation.
逐年被引趋势
暂无年度引用数据
关键指标
同类平均 = 1
同领域 · 同年份 · 同类型
Google Scholar 与 OpenAlex 的被引统计范围不同,数值存在差异属正常。
AI 辅助阅读
依据:摘要
回答优先基于摘要、文献信息与可获取全文;依据不足时会明确说明。
学术脉络
学科主题
计算机 / AIParallel Computing and Optimization Techniques
Embedded Systems Design Techniques · Interconnection Networks and Systems
参考文献 24
此处列出前 3 条