Not working in LM Studio

#1
by TreeLoys - opened

minicpm5-1b-Q8_0.gguf model, i say "hello" in chat, he no asked return. Any quest not worked.

Owner

Try again now. I've fixed the issue you might have been facing when running the models in LM Studio. Let me know if it worked!

Possible to disable thinking mode via chat template or system prompt?

minicpm5

it's still broken though

Tried everything, even downgraded the runtime to check if it might be llama cpp but nothing works

Owner

which quant you were using here Q4, Q8, or f16

Now. I have commented closed thread https://huggingface.co/openbmb/MiniCPM5-1B-GGUF/discussions/2 end i can run it. But need only llama.cpp last version build from git (for you OS and graphics card). The raw beta version LM Studio can't run it anyway (last or no). Use llama.cpp only from thread command. In CPU only mode have 33t/s on ryzen 5600.

Actually i found a solution of enabling LM Studio Engine Protocol in developer option working (llama.cpp b9334, lm studio 0.4.15 (build 1))

which quant you were using here Q4, Q8, or f16

I was using f16 though, but now working. thanks!

Sign up or log in to comment