Bone: Block-Affine Adaptation of Large Language Models

Kang, Jiale

Computer Science > Computation and Language

arXiv:2409.15371 (cs)

[Submitted on 19 Sep 2024 (v1), last revised 22 Nov 2024 (this version, v4)]

Title:Bone: Block-Affine Adaptation of Large Language Models

Authors:Jiale Kang

View PDF HTML (experimental)

Abstract:Low-Rank Adaptation (LoRA) has achieved remarkable training results by freezing the original weights and training only low-rank matrices, establishing itself as the predominant fine-tuning method for LLMs. In pursuit of performance closer to full-parameter training, a series of LoRA variants have emerged, such as LoRA+, PISSA, Olora, and LoRA-GA. This paper introduces a novel PEFT technique distinct from LoRA, called Block-Affine Adaptation (Bone). By dividing the original weights into multiple subspaces that share a single matrix for weight updates, Bone simplifies the process by requiring the trainable matrix to be initialized to zero, eliminating the need for complex initialization as in some LoRA variants. Compared to LoRA, Bone significantly reduces memory usage and achieves faster computation. Evaluation of both NLU and NLG tasks demonstrates that Bone substantially outperforms LoRA and its variants. Inspired by Pissa, we further proposed the ``Weight Guide'' theory to better utilize the information from the original weights. By integrating ``Weight Guide'' with Bone, we developed a new structure called Block-Affine Transformation (Bat), and ablation experiments confirmed the effectiveness of ``Weight Guide''.

Subjects:	Computation and Language (cs.CL); Artificial Intelligence (cs.AI)
Cite as:	arXiv:2409.15371 [cs.CL]
	(or arXiv:2409.15371v4 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.2409.15371

Submission history

From: Jiale Kang [view email]
[v1] Thu, 19 Sep 2024 10:26:42 UTC (3,326 KB)
[v2] Tue, 1 Oct 2024 10:00:49 UTC (2,011 KB)
[v3] Wed, 2 Oct 2024 07:38:02 UTC (2,001 KB)
[v4] Fri, 22 Nov 2024 10:40:35 UTC (2,055 KB)

Computer Science > Computation and Language

Title:Bone: Block-Affine Adaptation of Large Language Models

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:Bone: Block-Affine Adaptation of Large Language Models

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators