impossible → routine

Fine-tuning on one GPU

Fine-tuning a large language model into your own specialized version on a single machine.

Impossible

Microsoft researchers measure full fine-tuning of GPT-3 at 1.2 TB of GPU memory and call deploying independent fine-tuned copies prohibitively expensive.

1.2 TB of GPU memory

source

1 year, 11 months

Routine

QLoRA fine-tunes a 65-billion-parameter model on a single 48 GB GPU in 24 hours, reaching 99.3% of ChatGPT's level on the Vicuna benchmark.

one 48 GB GPU, 24 hours

source

All dated pairs