Question 9
You are using QLoRA to fine-tune a neural network. The weight matrix W has dimensions 1024 x 2048, and LoRA uses rank 8 matrices. Quantization is applied, reducing precision to 4bits. What is the memory required for storing the parameters introduced by QLoRA?
12,000 bytes
13,000 bytes
12,288 bytes
24,576 bytes