11,1 тыс подписчиков

QLoRA: Efficient Finetuning of Quantized LLMs

Model name Guanaco, outperforms all previous openly released models on the Vicuna benchmark, reaching 99.3% of the performance level of ChatGPT while only requiring 24 hours of finetuning on a single GPU.

QLoRA - эффективный метод файнтюнинга, который позволяет сократить использование памяти, чтобы произвести файнтюнинг модели с 65B параметрами на одном GPU 48 ГБ.

🖥 Github: https://github.com/artidoro/qlora

⏩ Paper: https://arxiv.org/abs/2305.14314

⭐️ Demo: https://huggingface.co/spaces/uwnlp/guanaco-playground-tgi

📌 Dataset: https://paperswithcode.com/dataset/ffhq

ai_machinelearning_big_data

QLoRA: Efficient Finetuning of Quantized LLMs Model name Guanaco, outperforms all previous openly released models on the Vicuna benchmark, reaching 99.

Около минуты

24 мая 2023