Performance and Efficiency Gains of NPU-Based Servers over GPUs for AI Model Inference
The exponential growth of AI applications has intensified the demand for efficient inference hardware capable of delivering low-latency, high-throughput, and energy-efficient performance. This study presents a systematic, empirical comparison of GPU- and NPU-based server platforms across key AI infe...
Na minha lista:
| Principais autores: | , |
|---|---|
| Format: | Artigo |
| Sprog: | Inglês |
| Udgivet: |
MDPI AG
2025-09-01
|
| Serier: | Systems |
| Fag: | |
| Online adgang: | https://www.mdpi.com/2079-8954/13/9/797 |
| Tags: |
Ingen Tags, Vær først til at tagge denne postø!
|
