Inconsistent reasoning

#1
by Stabhappy - opened

Hi, I notice that the model (after conversion to GGUF, running in llama.cpp) does not reason every turn like it should.
A simple "test" prompt only reasons 9/20 times, whereas a Unsloth quant of the base model reasons every turn.

I verified the chat template and darn near every other runtime setting is identical - this appears to be a weights issue (to the untrained eye).

Thanks

It indeed is because of higher divergence. I also haven’t run this very extensively so its high.

The model has been updated and its now much better at reasoning and has even less refusals

coder3101 changed discussion status to closed

Sign up or log in to comment