Running Agents 9 PEFT Method Comparison ⚖ 9 Explore PEFT method performance with interactive plots and tables
view article Article Transformers now runs llama.cpp quants +1 marcsun13, ArthurZ, lysandre • 3 days ago • 58
view reply We only included techniques that are implemented in PEFT, and those two are not implemented.
LoRA-FA: Memory-efficient Low-rank Adaptation for Large Language Models Fine-tuning Paper • 2308.03303 • Published Aug 7, 2023 • 4 • 2
LoRA+: Efficient Low Rank Adaptation of Large Models Paper • 2402.12354 • Published Feb 19, 2024 • 8 • 3
VeLoRA: Memory Efficient Training using Rank-1 Sub-Token Projections Paper • 2405.17991 • Published May 28, 2024 • 14 • 5
Robust and Efficient Fine-tuning of LLMs with Bayesian Reparameterization of Low-Rank Adaptation Paper • 2411.04358 • Published Aug 3, 2025 • 1
Block-Diagonal LoRA for Eliminating Communication Overhead in Tensor Parallel LoRA Serving Paper • 2510.23346 • Published Jan 6 • 1
LoRA-GA: Low-Rank Adaptation with Gradient Approximation Paper • 2407.05000 • Published Jul 6, 2024 • 1
LoftQ: LoRA-Fine-Tuning-Aware Quantization for Large Language Models Paper • 2310.08659 • Published Oct 12, 2023 • 30 • 5
A Rank Stabilization Scaling Factor for Fine-Tuning with LoRA Paper • 2312.03732 • Published Nov 28, 2023 • 13 • 1