Instructions to use ProCreations/Image-2.1-Calibrated-NVFP4 with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Diffusers
How to use ProCreations/Image-2.1-Calibrated-NVFP4 with Diffusers:
pip install -U diffusers transformers accelerate
import torch from diffusers import DiffusionPipeline # switch to "mps" for apple devices pipe = DiffusionPipeline.from_pretrained("ProCreations/Image-2.1-Calibrated-NVFP4", dtype=torch.bfloat16, device_map="cuda") prompt = "Astronaut in a jungle, cold color palette, muted colors, detailed, 8k" image = pipe(prompt).images[0] - Notebooks
- Google Colab
- Kaggle
- Local Apps Settings
- Draw Things
- DiffusionBee
Release calibrated Image2.1 NVFP4 transformer with dynamic scaling and BF16 rank correction, native SM120 runtime, quality evidence and real-time demo
1961af5 verified Download reports/benchmark.json from ProCreations/Image-2.1-Calibrated-NVFP4: direct link, hf CLI and curl.
- Browser
- Download file 1.16 kB
-
https://huggingface.co/ProCreations/Image-2.1-Calibrated-NVFP4/resolve/main/reports/benchmark.json
- Command line
-
hf download hf://ProCreations/Image-2.1-Calibrated-NVFP4/reports/benchmark.json
-
curl -L -o benchmark.json https://huggingface.co/ProCreations/Image-2.1-Calibrated-NVFP4/resolve/main/reports/benchmark.json
1.16 kB
| { | |
| "load_seconds": 2.297096138005145, | |
| "torch": "2.14.0+cu130", | |
| "gpu": "NVIDIA RTX PRO 6000 Blackwell Workstation Edition", | |
| "nvfp4_linears": 224, | |
| "steps": 40, | |
| "cfg": 1, | |
| "bf16_rank": 128, | |
| "attention_dtype": "bfloat16", | |
| "approximate_cache": false, | |
| "timing": { | |
| "1024": { | |
| "seconds": [ | |
| 4.547408219019417, | |
| 4.571850906999316, | |
| 4.592371508013457, | |
| 4.610020697989967, | |
| 4.624509044981096 | |
| ], | |
| "mean": 4.58923207540065, | |
| "warmup_seconds": 8.535866206977516, | |
| "peak_gb": 30.255306752 | |
| }, | |
| "2048": { | |
| "seconds": [ | |
| 32.58121982298326, | |
| 32.67262674000813, | |
| 32.71958359400742 | |
| ], | |
| "mean": 32.657810052332934, | |
| "warmup_seconds": 35.41050831298344, | |
| "peak_gb": 51.385951232 | |
| } | |
| }, | |
| "protocol": "CUDA synchronized; batch1; full40steps; includes encoder, denoising and VAE; excludes model load, resolution warmup and file writes. Prefix KV cache enabled. All large projections use either native NVFP4 with BF16 rank128 correction or explicitly listed calibrated FP8 safety layers. Compiled mode emulates intermediate precision casts." | |
| } |