QR رمز

Performance and Efficiency Gains of NPU-Based Servers over GPUs for AI Model Inference

The exponential growth of AI applications has intensified the demand for efficient inference hardware capable of delivering low-latency, high-throughput, and energy-efficient performance. This study presents a systematic, empirical comparison of GPU- and NPU-based server platforms across key AI infe...

وصف كامل

محفوظ في:
التفاصيل البيبلوغرافية
المؤلفون الرئيسيون: Youngpyo Hong, Dongsoo Kim
التنسيق: Artigo
اللغة:Inglês
منشور في: MDPI AG 2025-09-01
سلاسل:Systems
الموضوعات:
الوصول للمادة أونلاين:https://www.mdpi.com/2079-8954/13/9/797
الوسوم: إضافة وسم
لا توجد وسوم, كن أول من يضع وسما على هذه التسجيلة!