Can't be quantized

#2
by saipangon - opened

How we quantize this into GGUF?

How we quantize this into GGUF?

you need to use some llm to put patches to llama cpp to make gguf -- use any opencopde free model -- give the downlaoed weights and latest llama cpp source -- and ask to make bf16 gguf for this

Sign up or log in to comment