Image21-INT4 / CHANGES.md
ixim's picture
Release verified Image21-INT4 conversion
9116984 verified
|
Raw History Blame
653 Bytes

Modifications

Modified by ixim / iximbox: eligible linear weights converted from Qwen-Image-2.1 to SDNQ UINT4 with SVD rank 32. Built with Qwen. Non-commercial research/evaluation under the accompanying Qwen Research License.

  • Converted eligible transformer and text-encoder linear layers to SDNQ UINT4.
  • Stored a rank-32 SVD residual of the quantization error with the weights.
  • Left the requested sensitive projections, normalization, embeddings, vision tower, output head and VAE in floating point.
  • Did not use a calibration set or fine-tuning.
  • Quantized matmul is off so CUDA and Apple Silicon use the same eager dequantization.