Back to library

AI / Technology

LoRALow-Rank Adaptation of Large Language Models

Three key questions about this paper

What problem does LoRA: Low-Rank Adaptation of Large Language Models address?

Authors: Edward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen Source: arXiv:2106.09685, version 2 (16 October 2021) · PDF Reading note: “Low rank” here describes the shape of the learned weight update, not a smaller base model.

What evidence supports the main claim in LoRA: Low-Rank Adaptation of Large Language Models?

The evaluation spans RoBERTa and DeBERTa on GLUE, GPT-2 on generation tasks, and GPT-3 175B on WikiSQL, MultiNLI, and SAMSum. Baselines include full fine-tuning, bias-only tuning, adapters, and prefix-based methods. All experiments used NVIDIA V100 GPUs except the separate GPT-2 adapter-latency study, which used a Quadro RTX 8000. [Paper §5, pp. 5–8; Table 1, p. 4]

What limitation should readers know about LoRA: Low-Rank Adaptation of Large Language Models?

The evaluation spans RoBERTa and DeBERTa on GLUE, GPT-2 on generation tasks, and GPT-3 175B on WikiSQL, MultiNLI, and SAMSum. Baselines include full fine-tuning, bias-only tuning, adapters, and prefix-based methods. All experiments used NVIDIA V100 GPUs except the separate GPT-2 adapter-latency study, which used a Quadro RTX 8000. [Paper §5, pp. 5–8; Table 1, p. 4]

2 new free reports left todaySubscribe to Pro for unlimited reading and 10 new paper explanations each month.Upgrade to Pro