2026· Interdisciplinary Journal of Computing & AI· 0 citations· 19 references
TL;DR
Simulation-based practice to improve the energy efficiency of Intel Arc GPUs and the Intel CPUs to overcome the issue of power inefficiency, workload imbalance, and thermal limitations indicates that the Intel Arc GPUs used less power compared to CPUs with similar tasks, which makes the argument of their energy efficiency advantage in environments with limited energy.
Abstract
The growing popularity of High-Performance Computing (HPC), artificial intelligence (AI) and sophisticated graphics display has rendered the management of GPS power as an important design factor. Proposed methodology presented in this paper, is a simulation-based practice to improve the energy efficiency of Intel Arc™ GPUs and the Intel CPUs to overcome the issue of power inefficiency, workload imbalance, and thermal limitations. The diagrammatic analysis of MATLAB/Simulink model is used to study the dynamic power behavior of system components under different load conditions. The Intel Arc A770, A750, and B580 GPUs, the Intel Core i7-14700K processor, and a high-voltage Switched-Mode Power Supply (SMPS) are implemented in the model and make it possible to simulate infrastructure realistically. Dynamic Voltage and Frequency Scaling (DVFS), idle power gating and workload-aware scheduling were all applied using the control systems and power electronics toolboxes in MATLAB. Validation of the experiment was conducted by real-time telemetry logging and Python based analysis of power, usage, temperature and frequency metrics of gaming, AI and compute workloads. Findings indicate that, in high-intensity tasks, the Intel Arc GPUs used less power compared to CPUs with similar tasks, which makes the argument of their energy efficiency advantage in environments with limited energy. The model also captures the thermal feedback and the voltage control in the SMPS and makes it stable under varying loads. Researchers and engineers can use this open-source and reproducible tool to obtain actionable insights for micro-architecture and system design of power-efficient high performance computing systems that are adaptable for innovation to concurrent hardware technologies and emerging, sustainability-driven demands.
A unique combination of Posit arithmetic, iterative division, SIMD parallelisation, and pipelining optimisation into a single framework that addresses the shortcomings of IEEE-754 floating-point arithmetic and can be used effectively in high-performance computing systems, embedded processors, and AI accelerators.
K. Pande, P. Karule· African Journal Of Applied R...· 0 citations
A graphics processor (GPU) is a dedicated processing unit that handles all graphics-related computation in a given system. With the growth of graphics-based applications in domains such as embedded systems, gaming, and image processing, studying such systems has become more important. We propose to develop a simulator...
This work evaluates whether Model FLOPs Utilization (MFU) can serve as a portable, software-defined predictor of GPU power for LLMs, finding that a linear MFU-based power model fits every tested GPU as long as the workload is compute-bound, as in production LLM training.
DiffPower translates design netlists into a PDK-agnostic bytecode representation, enabling analytical gradient computation via reverse-mode automatic differentiation, achieving up to a speedup over single-threaded CPU propagation on the largest evaluated design, with the GPU advantage growing with design scale.
Isaac Jacobson, Zhengjie Zhao, Rashmi Mehrotra et al.· 0 citations
MEPOWER is proposed, a flexible, model-based approach to exposing compute/data movement imbalance that characterizes the fine-grained memory behavior of parallel workloads that demonstrates a reduction in EDP on a range of HPC benchmarks with minimal impact on execution time when compared to the standard OS/hardware-ma...
Nanda Velugoti, Joseph Manzano, Andrés Márquez et al.· 0 citations
A benchmarking study of Python-Rust interoperability reveals that the Rust enhanced versions systematically execute faster than pure Pythoncode- up to three times faster in some cases.
Srikant Singh, R.pradeep Raj· International Journal For Mu...· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.