QRコード

Efficient Fine-Tuning of Large Language Models via a Low-Rank Gradient Estimator

In this paper, we present a Low-Rank Gradient Estimator (LoGE) to accelerate the finetune-time computation of transformers, especially large language models (LLMs). Unlike Parameter-Efficient Fine-Tuning (PEFT) methods, which primarily aim to minimize the number of fine-tuning parameters, LoGE also...

詳細記述

保存先:
書誌詳細
主要な著者: Luoming Zhang, Zhenyu Lou, Yangwei Ying, Cheng Yang, Hong Zhou
フォーマット: Artigo
言語:Inglês
出版事項: MDPI AG 2024-12-01
シリーズ:Applied Sciences
主題:
オンライン・アクセス:https://www.mdpi.com/2076-3417/15/1/82
タグ: タグ追加
タグなし, このレコードへの初めてのタグを付けませんか!