Before you start
Hugging Face link
https://huggingface.co/Accio-Lab/occamy-1.0-NVFP4
Is the model architecture already supported
Yes, but this checkpoint or quantization does not load
Is the quantization already supported
Yes, but this checkpoint's weight format does not load
What happens when you load it
>>>>>>>> PASTE THE FULL SERVER LOG HERE <<<<<<<<
(copied with the "Copy server log" button in the failure dialog, or Logs -> Server status -> Copy)
WeightLoadError: ValueError: model.layers.0.mlp.gate.weight is torch.bfloat16 but the checkpoint's quant config declares model.layers.0.mlp.gate QuantScheme(nvfp4, WeightDesc(elem='e2m1', group=(1, 16
Anything else
Desktop 0.2.0-beta.22; engine 0.1.3+gc8ed699cb; NVIDIA GeForce RTX 5080, 15.9 GiB, driver 610.74; 13th Gen Intel Core i7-13700KF, 31.8 GiB
Before you start
Hugging Face link
https://huggingface.co/Accio-Lab/occamy-1.0-NVFP4
Is the model architecture already supported
Yes, but this checkpoint or quantization does not load
Is the quantization already supported
Yes, but this checkpoint's weight format does not load
What happens when you load it
Anything else
Desktop 0.2.0-beta.22; engine 0.1.3+gc8ed699cb; NVIDIA GeForce RTX 5080, 15.9 GiB, driver 610.74; 13th Gen Intel Core i7-13700KF, 31.8 GiB