Models not loading into RAM

corvus@lemmy.ml · 2 months ago

Models not loading into RAM

Possibly linux · edit-2 2 months ago

Did you try your CPU?

Also try Deepseek 14b. It will be much faster.

corvus@lemmy.ml · edit-2 2 months ago

Yes, gpt4all runs it in cpu mode, the gpu option does not appear in the drop-down menu, which means the gpu it’s not supported or there is an error. I’m trying to run the models with the SyCL backend implemented in llama.cpp that performs specific optimizations for cpu+gpu with the Intel DPC++/C++ Compiler and the OneAPI Toolkit.

Also try Deepseek 14b. It will be much faster.

ok, I’ll test it out.

Possibly linux · 2 months ago

Why don’t you just use ollama?

corvus@lemmy.ml · 2 months ago

I don’t like intermediaries ;) Fortunately I compiled llama.cpp with the Vulkan backend and everything went smooth and now I have the option to offload to the GPU. Now I will test performance CPU vs CPU+GPU. Downloaded deepseek 14b and is really good, the best I could run so far in my limited hardware.