Configuration Parsing Warning:In config.json: "quantization_config.modules_to_not_convert" must be an array

GLM-4-9B-0414_FP8

Introduction

FP8 weights quantization of https://huggingface.co/zai-org/GLM-4-9B-0414

Downloads last month
12
Safetensors
Model size
9B params
Tensor type
F32
·
F8_E4M3
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for bunbohue/GLM-4-9B-0414-FP8

Quantized
(25)
this model