Qwen3.5 (quantized)

Quantized versions of text_encoders/qwen3.5_4b_bf16.safetensors from Comfy-Org/Qwen3.5. Nothing else is changed.

File Format
qwen3.5_4b_int8_convrot int8 ConvRot
qwen3.5_4b_fp8_scaled fp8 e4m3fn

The int8 ConvRot file is probably the better choice: it stays closer to bf16.

Only the language-model linear layers are quantized, in ComfyUI's comfy_quant format (int8: ConvRot group 256 with per-row scales; fp8: per-tensor scale). Embeddings, norms, the vision encoder and the MTP head stay bf16.

Put the file in ComfyUI/models/text_encoders.

Downloads last month
62
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for nomadoor/Qwen3.5

Finetuned
Qwen/Qwen3.5-4B
Finetuned
(850)
this model