Related Experiment Video
Updated: Jun 26, 2025

Assessing the Multiple Dimensions of Engagement to Characterize Learning: A Neurophysiological Perspective
Published on: July 1, 2015
Exploring memory synchronization and performance considerations for FPGA platform using the high-abstracted OpenCL
Abedalmuhdi Almomany1,2, Amin Jarrah2, Muhammed Sutcu3
1Department of Electrical & Computer Engineering, Gulf University for Science & Technology, Kuwait, Kuwait.
Abstract:
A key benefit of the Open Computing Language (OpenCL) software framework is its capability to operate across diverse architectures. Field programmable gate arrays (FPGAs) are a high-speed computing architecture used for computation acceleration. This study investigates the impact of memory access time on overall performance in general FPGA computing environments through the creation of eight benchmarks within the OpenCL framework. The developed benchmarks capture a range of memory access behaviors, and they play a crucial role in assessing the performance of spinning and sleeping on FPGA-based architectures. The results obtained guide the formulation of new implementations and contribute to defining an abstraction of FPGAs. This abstraction is then utilized to create tailored implementations of primitives that are well-suited for this platform. While other research endeavors concentrate on creating benchmarks with the Compute Unified Device Architecture (CUDA) to scrutinize the memory systems across diverse GPU architectures and propose recommendations for future generations of GPU computation platforms, this study delves into the memory system analysis for the broader FPGA computing platform. It achieves this by employing the highly abstracted OpenCL framework, exploring various data workload characteristics, and experimentally delineating the appropriate implementation of primitives that can seamlessly integrate into a design tailored for the FPGA computing platform. Additionally, the results underscore the efficacy of employing a task-parallel model to mitigate the need for high-cost synchronization mechanisms in designs constructed on general FPGA computing platforms.
More Related Videos
08:09Multifunctional Setup for Studying Human Motor Control Using Transcranial Magnetic Stimulation, Electromyography, Motion Capture, and Virtual Reality
Published on: September 3, 2015
08:07Assembly and Characterization of Biomolecular Memristors Consisting of Ion Channel-doped Lipid Membranes
Published on: March 9, 2019
Related Concept Videos
Clamper Circuit
Within this circuit, the diode's orientation prompts the capacitor to charge up to the level of the most negative peak of the input signal. Upon reaching this state, the diode ceases to...
Non-ohmic Devices
Consider a simple circuit consisting of a battery, a diode, and a resistor. A...
Sleep-Wake Cycles
NREM Sleep
NREM sleep comprises four progressive stages that seamlessly merge:
Buffers: Buffer Capacity
In the graph, pH is plotted as a function of the number of moles of base (Cb) added to a weak...
Linear time-invariant Systems
The input-output behavior of an LTI system can be fully defined by its response to an impulsive excitation at its input. Once this impulse response is known, the system's reaction to any other input can be...
Fast Decoupled and DC Powerflow