I'm quite new to LLM world. I'm brazilian and decided to start with Cabrita. I managed to deal with several errors but i am stuck with this:
"Some modules are dispatched on the CPU or the disk. Make sure you have enough GPU RAM to fit
the quantized model. If you have set a value for max_memory you should increase that. To have
an idea of the modules that are set on the CPU or RAM you can print model.hf_device_map."
I tried several solutions but none work. Is it a memory problem?
I have a MacBook Pro i7 16Gb (2018)
I'm quite new to LLM world. I'm brazilian and decided to start with Cabrita. I managed to deal with several errors but i am stuck with this:
"Some modules are dispatched on the CPU or the disk. Make sure you have enough GPU RAM to fit
the quantized model. If you have set a value for
max_memoryyou should increase that. To havean idea of the modules that are set on the CPU or RAM you can print model.hf_device_map."
I tried several solutions but none work. Is it a memory problem?
I have a MacBook Pro i7 16Gb (2018)