Pruning Adapters with Lottery Ticket
Massively pre-trained transformer models such as BERT have gained great success in many downstream NLP tasks. However, they are computationally expensive to fine-tune, slow for inference, and have large storage requirements. So, transfer learning with adapter modules has been introduced and has beco...
Gespeichert in:
| Hauptverfasser: | , |
|---|---|
| Format: | Artigo |
| Sprache: | Inglês |
| Veröffentlicht: |
MDPI AG
2022-02-01
|
| Schriftenreihe: | Algorithms |
| Schlagworte: | |
| Online-Zugang: | https://www.mdpi.com/1999-4893/15/2/63 |
| Tags: |
Keine Tags, Fügen Sie das erste Tag hinzu!
|
