Sign In

Qwen Image 2.1 MXFP8

Updated: Sep 29, 2026

base model

Download

1 variant available

mxfp8 SafeTensor

qwen_image_v21_quant_mxfp8.safetensors

Block-scaled 8-bit, near-FP16 quality • 6.9 GB

Verified:

Type
Checkpoint Trained
Stats

38

Reviews
Published

Sep 29, 2026

Base Model

Qwen 2.1

Hash
AutoV2
C3CB51E3C9
default creator card background decoration
Reactions - 188

188

Followers - 12

12

Downloads - 689

689

Qwen is licensed under the Qwen RESEARCH LICENSE AGREEMENT, Copyright (c) 2026 Hangzhou Tongyi Laboratory Technology Co., Ltd. All Rights Reserved.

ComfyUI_temp_mihni_00009_.png

# Qwen Image 2.1 [MXFP8] 🚀

This is the optimized MXFP8 quantized version of the powerful Qwen Image 2.1 model. The goal of this upload is to bring the excellent visual generation and prompt comprehension capabilities of Qwen 2.1 to setups with limited resources, drastically reducing VRAM consumption with no noticeable loss in quality.

## ✨ Main Highlights

* MXFP8 Efficiency: Utilizes the Microscaling FP8 format to compress the model. This means it takes up almost half the disk space and VRAM compared to FP16/BF16 versions, while maintaining detail precision and structural fidelity.

* Accelerated Inference: Modern GPUs (RTX 5000/4000/3000 series and equivalents) benefit greatly from FP8 compute, resulting in significantly faster generation and processing times.

* Accessibility: Perfect for running locally on graphics cards with lower VRAM, allowing for heavier workflows or higher resolutions that would normally cause an Out of Memory (OOM) error on the base model.