MicroBERT: Distilling MoE-Based Knowledge from BERT into a Lighter Model
Natural language-processing tasks have been improved greatly by large language models (LLMs). However, numerous parameters make their execution computationally expensive and difficult on resource-constrained devices. For this problem, as well as maintaining accuracy, some techniques such as distilla...
Bewaard in:
| Hoofdauteurs: | , , , , |
|---|---|
| Formaat: | Artigo |
| Taal: | Inglês |
| Gepubliceerd in: |
MDPI AG
2024-07-01
|
| Reeks: | Applied Sciences |
| Onderwerpen: | |
| Online toegang: | https://www.mdpi.com/2076-3417/14/14/6171 |
| Tags: |
Geen labels, Wees de eerste die dit record labelt!
|
