Performance and Efficiency Gains of NPU-Based Servers over GPUs for AI Model Inference
The exponential growth of AI applications has intensified the demand for efficient inference hardware capable of delivering low-latency, high-throughput, and energy-efficient performance. This study presents a systematic, empirical comparison of GPU- and NPU-based server platforms across key AI infe...
محفوظ في:
| المؤلفون الرئيسيون: | , |
|---|---|
| التنسيق: | Artigo |
| اللغة: | Inglês |
| منشور في: |
MDPI AG
2025-09-01
|
| سلاسل: | Systems |
| الموضوعات: | |
| الوصول للمادة أونلاين: | https://www.mdpi.com/2079-8954/13/9/797 |
| الوسوم: |
لا توجد وسوم, كن أول من يضع وسما على هذه التسجيلة!
|
