Has anyone tried this

#2
by dpe1 - opened

I downloaded this model using hf CLI, because unsloth is notoriously bad at downloading models, but now for some reason Unsloth doesn't accept my model and wants to download it again, which it fails to do of course, so I am stuck with not being able to test this for now

wdym? unsloth is annoyinglyy good at managing models. maybe because this is not a GGUF file :(

wdym? unsloth is annoyinglyy good at managing models. maybe because this is not a GGUF file :(

well I can't load this model in unsloth, it just doesn't do anything, keeps being stuck at loading weights, in the mean time I loaded JohnRoger/VibeThinker-3B-Q8_0-GGUF but I think there is an issue with that model, because it starts reasoning about its own chat template, so I think the chat template is wrong in it

Ahh, sometimes it happens when the community duct type things. letm guess, the model start doing something like:

<think>
the user asked [question] hmm... what does /think mean? maybe the user was trying to indicate thinking traces?

</think>

bcs that is the EXACT thing whch happened to me with an older depsek disill model

Sign up or log in to comment