impossible → routine
Fine-tuning on one GPU
Fine-tuning a large language model into your own specialized version on a single machine.
Impossible
Microsoft researchers measure full fine-tuning of GPT-3 at 1.2 TB of GPU memory and call deploying independent fine-tuned copies prohibitively expensive.
1.2 TB of GPU memory
source1 year, 11 months
Routine
QLoRA fine-tunes a 65-billion-parameter model on a single 48 GB GPU in 24 hours, reaching 99.3% of ChatGPT's level on the Vicuna benchmark.
one 48 GB GPU, 24 hours
source