Wrongly made GGUF

#1
by Nerdsking - opened

Check this: https://github.com/ggml-org/llama.cpp/pull/24260
The llama already released an official version (https://github.com/ggml-org/llama.cpp/releases/tag/b9626) for the CohereLabs.command-a-plus-05-2026-bf16.
This GGUF however will cause an error, because it is using "cohere2-moe", while the correct would be "cohere2moe". Better check this.

DevQuasar org

thanks for the notification.
This is an experimental quant as I've stated on the model card.
Will re-quantize the model in the upcoming days

DevQuasar org

done

csabakecskemeti changed discussion status to closed

Sign up or log in to comment