Código QR (código de barras bidimensional)

Performance and Efficiency Gains of NPU-Based Servers over GPUs for AI Model Inference

The exponential growth of AI applications has intensified the demand for efficient inference hardware capable of delivering low-latency, high-throughput, and energy-efficient performance. This study presents a systematic, empirical comparison of GPU- and NPU-based server platforms across key AI infe...

Fuld beskrivelse

Na minha lista:
Bibliografiske detaljer
Principais autores: Youngpyo Hong, Dongsoo Kim
Format: Artigo
Sprog:Inglês
Udgivet: MDPI AG 2025-09-01
Serier:Systems
Fag:
Online adgang:https://www.mdpi.com/2079-8954/13/9/797
Tags: Tilføj Tag
Ingen Tags, Vær først til at tagge denne postø!