Couldn't load on 64 GB P14s

#1
by leouon - opened

Thank you for this quantization it was really interesting to see a new model like this from this lineage of training but the Q4_K_M quant didn't load in 64GB llama.cpp. I'll try on another machine too.

Hm, that's interesting. I can fit Q6_K in 64GB just fine, though that's purely ram only. Can you share your log? Maybe I could figure it out.

By P14 btw, do you mean a Lenovo Thinkpad P14?

Yes that's right. Sorry I haven't got back to you. I was able to get the E4B working, I will be doing a little more with some other things and putting this a bit on the backburner. I was able to get a merged version of the 26A4B working so it's ok. Thank you for the quantizations wishing you well.

leouon changed discussion status to closed

Sign up or log in to comment