Eliovp commited on
Commit
ddc4f70
路
verified 路
1 Parent(s): 680aa3e

Release v1.0.1 with persistent Docker cache setup

Browse files
Files changed (4) hide show
  1. BENCHMARKS.md +1 -1
  2. README.md +8 -5
  3. SHA256SUMS +3 -3
  4. paiton-hub.json +12 -12
BENCHMARKS.md CHANGED
@@ -90,6 +90,6 @@ other generation services, then run:
90
 
91
  [Every measured timing](benchmark-data/raw-timings.csv) and
92
  [settings, memory, versions and quality](benchmark-data/results.json) are included.
93
- The [article and evidence bundle](https://github.com/Eliovp-BV/paiton-vllm-plugin/releases/tag/paiton-flux2-klein-gfx1201-v1.0.0)
94
  provides charts, full-resolution quality pairs, example media, startup details
95
  and the complete evidence archive.
 
90
 
91
  [Every measured timing](benchmark-data/raw-timings.csv) and
92
  [settings, memory, versions and quality](benchmark-data/results.json) are included.
93
+ The [article and evidence bundle](https://github.com/Eliovp-BV/paiton-vllm-plugin/releases/tag/paiton-flux2-klein-gfx1201-v1.0.1)
94
  provides charts, full-resolution quality pairs, example media, startup details
95
  and the complete evidence archive.
README.md CHANGED
@@ -20,7 +20,7 @@ tags:
20
 
21
  Generate photographs, product concepts and illustrations locally with a ready-to-run ComfyUI workflow. The qualified Paiton pipeline averages **1.054 seconds per 1024 脳 1024 image** and uses just **12.9 GiB peak Torch allocation** on one AMD Radeon AI PRO R9700.
22
 
23
- **Release:** v1.0.0 路 **GPU:** R9700, `gfx1201`, 32 GB 路 **Steps:** 4 路 **Batch:** 1 路 **Text-to-image**
24
 
25
  This repository contains compiled runtime artifacts, manifests, licenses and example media. Model weights download separately from the original checkpoint publisher and remain in your local cache. The containers already include these artifacts, so no separate `.so` download is needed to start. The private Paiton compiler is not included or required.
26
 
@@ -29,18 +29,20 @@ This repository contains compiled runtime artifacts, manifests, licenses and exa
29
  You need Linux x86-64, Docker Engine with the Compose plugin, and working AMD GPU device access through `/dev/kfd` and `/dev/dri`.
30
 
31
  ```bash
32
- git clone --depth 1 --branch paiton-flux2-klein-gfx1201-v1.0.0 https://github.com/Eliovp-BV/paiton-vllm-plugin.git && cd paiton-vllm-plugin && ./models/FLUX.2-klein/launch.sh
33
  ```
34
 
35
  Open [ComfyUI on localhost](http://127.0.0.1:8188/?paiton=1). The first visit opens the included workflow. Enter a prompt, choose **Paiton** or **Stock (Diffusers)**, and click **Run**. Preview and Save Image work normally. Images go to `paiton-images/`.
36
 
37
  ![ComfyUI with the included Paiton workflow](assets/comfyui-workflow.png)
38
 
 
 
39
  The helper pulls three public containers, downloads the pinned 5.46 GB checkpoint and prepares about 12 GB of runtime tensors once. Source weights are preserved. Downloads, prepared tensors, compilation caches and UI settings persist. No Hugging Face token or paid service is required.
40
 
41
  The first image loads and compiles the selected engine and can take several minutes. Later images reuse it. Switching engines unloads the previous model and incurs setup again. For comparisons, keep the prompt and seed fixed and choose **fixed** in the seed control.
42
 
43
- From `models/FLUX.2-klein`, run `./launch.sh --stop` to stop, `./launch.sh --logs` for logs, or `./launch.sh --ui simple` for a smaller prompt-and-image interface. Stopping preserves caches and outputs. The [complete model guide](https://github.com/Eliovp-BV/paiton-vllm-plugin/tree/paiton-flux2-klein-gfx1201-v1.0.0/models/FLUX.2-klein) covers terminal generation and using the node in an existing ComfyUI installation. This image package uses Diffusers; the same central repository also contains our vLLM text releases.
44
 
45
  ## 12.9 GiB peak Torch allocation
46
 
@@ -98,7 +100,7 @@ The pinned runtime includes Torch 2.12.0 with ROCm 7.14, Diffusers 0.40.0 and Tr
98
  The quick start already contains these exact files. For a separate, versioned copy:
99
 
100
  ```bash
101
- hf download EliovpAI/FLUX.2-klein-4B-Paiton-RDNA4 --revision v1.0.0 --include 'artifacts/*' --local-dir ./paiton-flux2-hf
102
  ```
103
 
104
  Keep all six files in `artifacts/` together. They contain three shared libraries and their hash-checked compatibility manifests. They are loaded by the Paiton image runtime, not by `DiffusionPipeline.from_pretrained` as model weights.
@@ -110,9 +112,10 @@ mkdir -p paiton-images
110
  docker run --rm --device /dev/kfd --device /dev/dri \
111
  --group-add "$(stat -c '%g' /dev/kfd)" --shm-size 2g \
112
  -v paiton-flux2-cache:/models \
 
113
  -v "$PWD/paiton-flux2-hf/artifacts:/opt/paiton/image/artifacts:ro" \
114
  -v "$PWD/paiton-images:/outputs" \
115
- ghcr.io/eliovp/paiton-vllm-plugin:flux2-klein-rdna4-v1.0.0 \
116
  generate --prompt 'A fox in a woodland at sunrise, wildlife photograph' --seed 42
117
  ```
118
 
 
20
 
21
  Generate photographs, product concepts and illustrations locally with a ready-to-run ComfyUI workflow. The qualified Paiton pipeline averages **1.054 seconds per 1024 脳 1024 image** and uses just **12.9 GiB peak Torch allocation** on one AMD Radeon AI PRO R9700.
22
 
23
+ **Release:** v1.0.1 路 **GPU:** R9700, `gfx1201`, 32 GB 路 **Steps:** 4 路 **Batch:** 1 路 **Text-to-image**
24
 
25
  This repository contains compiled runtime artifacts, manifests, licenses and example media. Model weights download separately from the original checkpoint publisher and remain in your local cache. The containers already include these artifacts, so no separate `.so` download is needed to start. The private Paiton compiler is not included or required.
26
 
 
29
  You need Linux x86-64, Docker Engine with the Compose plugin, and working AMD GPU device access through `/dev/kfd` and `/dev/dri`.
30
 
31
  ```bash
32
+ git clone --depth 1 --branch paiton-flux2-klein-gfx1201-v1.0.1 https://github.com/Eliovp-BV/paiton-vllm-plugin.git && cd paiton-vllm-plugin && ./models/FLUX.2-klein/launch.sh
33
  ```
34
 
35
  Open [ComfyUI on localhost](http://127.0.0.1:8188/?paiton=1). The first visit opens the included workflow. Enter a prompt, choose **Paiton** or **Stock (Diffusers)**, and click **Run**. Preview and Save Image work normally. Images go to `paiton-images/`.
36
 
37
  ![ComfyUI with the included Paiton workflow](assets/comfyui-workflow.png)
38
 
39
+ The prepared tensors use the `paiton-flux2-cache` volume; source weights and compiler caches use `paiton-flux2-cache-runtime`. Both persist across complete stop/start cycles.
40
+
41
  The helper pulls three public containers, downloads the pinned 5.46 GB checkpoint and prepares about 12 GB of runtime tensors once. Source weights are preserved. Downloads, prepared tensors, compilation caches and UI settings persist. No Hugging Face token or paid service is required.
42
 
43
  The first image loads and compiles the selected engine and can take several minutes. Later images reuse it. Switching engines unloads the previous model and incurs setup again. For comparisons, keep the prompt and seed fixed and choose **fixed** in the seed control.
44
 
45
+ From `models/FLUX.2-klein`, run `./launch.sh --stop` to stop, `./launch.sh --logs` for logs, or `./launch.sh --ui simple` for a smaller prompt-and-image interface. Stopping preserves caches and outputs. The [complete model guide](https://github.com/Eliovp-BV/paiton-vllm-plugin/tree/paiton-flux2-klein-gfx1201-v1.0.1/models/FLUX.2-klein) covers terminal generation and using the node in an existing ComfyUI installation. This image package uses Diffusers; the same central repository also contains our vLLM text releases.
46
 
47
  ## 12.9 GiB peak Torch allocation
48
 
 
100
  The quick start already contains these exact files. For a separate, versioned copy:
101
 
102
  ```bash
103
+ hf download EliovpAI/FLUX.2-klein-4B-Paiton-RDNA4 --revision v1.0.1 --include 'artifacts/*' --local-dir ./paiton-flux2-hf
104
  ```
105
 
106
  Keep all six files in `artifacts/` together. They contain three shared libraries and their hash-checked compatibility manifests. They are loaded by the Paiton image runtime, not by `DiffusionPipeline.from_pretrained` as model weights.
 
112
  docker run --rm --device /dev/kfd --device /dev/dri \
113
  --group-add "$(stat -c '%g' /dev/kfd)" --shm-size 2g \
114
  -v paiton-flux2-cache:/models \
115
+ -v paiton-flux2-cache-runtime:/models/cache \
116
  -v "$PWD/paiton-flux2-hf/artifacts:/opt/paiton/image/artifacts:ro" \
117
  -v "$PWD/paiton-images:/outputs" \
118
+ ghcr.io/eliovp/paiton-vllm-plugin:flux2-klein-rdna4-v1.0.1 \
119
  generate --prompt 'A fox in a woodland at sunrise, wildlife photograph' --seed 42
120
  ```
121
 
SHA256SUMS CHANGED
@@ -1,8 +1,8 @@
1
- d032d165b20f77321938e6371d91345e67c3f3328f815ad3d4da1a58b35bf984 BENCHMARKS.md
2
  c71d239df91726fc519c6eb72d318ec65820627232b2f796219e87dcf35d0ab4 LICENSE
3
  92640fb97222fd0a698ff28ce0c3782c172623f8d6c609b557636a80f28fb946 LICENSES/Triton-MIT.txt
4
  3419cabb33500032ddf1adbb80d5e525adff6feeb0b42b6fe799d850bd860a06 NOTICE
5
- d5e5824b4a3d26005e45e631448f659bef20a226d09f8023c7a72e86d781d103 README.md
6
  dde1f76fa30cceb4487b492b61332b90ebb2b478b15a733fba12af3aaa2068b0 THIRD_PARTY_NOTICES.md
7
  693063398478a32235f27ffc9b324811c23d588c02f1283d148499d60da05484 artifacts/flux2_klein_decoder_gfx1201.manifest.json
8
  456f278c21a25445edd1d1d91196cd63f0f5ac7eb9bc0b146d503c526ea316cf artifacts/flux2_klein_decoder_gfx1201.so
@@ -19,4 +19,4 @@ d957298bc2fb8a4f4870411968bc4319c169857f2d167d75508fba0766474b27 assets/product
19
  531be4d5692bb4867571df3e2fc40225a6c0b78284834b0daa449071be908adb benchmark-data/comfyui-validation.json
20
  aa618807100cf1d5d3affbeffaf161314b933e8fc29933cf0f0a6b915a5be80d benchmark-data/raw-timings.csv
21
  fbf4587152044ebbef956df16b46b9b22c9dec65940c881e666749d5721aad86 benchmark-data/results.json
22
- cd848e6f0d2430d9f8392ce98df05b25021bf8ca9eedefdc6eed0ae89b7dc6db paiton-hub.json
 
1
+ 7ee29787618ca4b93981fb0214af433787334c331c16b64050e4b054e2b01de9 BENCHMARKS.md
2
  c71d239df91726fc519c6eb72d318ec65820627232b2f796219e87dcf35d0ab4 LICENSE
3
  92640fb97222fd0a698ff28ce0c3782c172623f8d6c609b557636a80f28fb946 LICENSES/Triton-MIT.txt
4
  3419cabb33500032ddf1adbb80d5e525adff6feeb0b42b6fe799d850bd860a06 NOTICE
5
+ 51e4bf6eca8277ac9811da7ff6f742c1f0247182c3923b41c31f7f700dc2e11a README.md
6
  dde1f76fa30cceb4487b492b61332b90ebb2b478b15a733fba12af3aaa2068b0 THIRD_PARTY_NOTICES.md
7
  693063398478a32235f27ffc9b324811c23d588c02f1283d148499d60da05484 artifacts/flux2_klein_decoder_gfx1201.manifest.json
8
  456f278c21a25445edd1d1d91196cd63f0f5ac7eb9bc0b146d503c526ea316cf artifacts/flux2_klein_decoder_gfx1201.so
 
19
  531be4d5692bb4867571df3e2fc40225a6c0b78284834b0daa449071be908adb benchmark-data/comfyui-validation.json
20
  aa618807100cf1d5d3affbeffaf161314b933e8fc29933cf0f0a6b915a5be80d benchmark-data/raw-timings.csv
21
  fbf4587152044ebbef956df16b46b9b22c9dec65940c881e666749d5721aad86 benchmark-data/results.json
22
+ f6a192228680ae571e170a8e30ffd613471e20ec670f0f23f5b0306052a87d8f paiton-hub.json
paiton-hub.json CHANGED
@@ -1,12 +1,12 @@
1
  {
2
  "format_version": 1,
3
  "repository": "EliovpAI/FLUX.2-klein-4B-Paiton-RDNA4",
4
- "version": "v1.0.0",
5
  "content": "Compiled runtime artifacts, manifests, notices, benchmark data and example media. No checkpoint weights.",
6
  "runtime_source": {
7
  "repository": "https://github.com/Eliovp-BV/paiton-vllm-plugin",
8
- "revision": "0e20f1db97050f9eb83af0158f6c6d5bd329e0a1",
9
- "tag": "paiton-flux2-klein-gfx1201-v1.0.0",
10
  "path": "models/FLUX.2-klein"
11
  },
12
  "source_checkpoint": {
@@ -16,19 +16,19 @@
16
  },
17
  "containers": {
18
  "klein": {
19
- "tag": "ghcr.io/eliovp/paiton-vllm-plugin:flux2-klein-rdna4-v1.0.0",
20
- "digest": "sha256:8866c8dd29d7dfbba81156a83af75229ffc9bc05a074833bc516d8da79ebd6a8",
21
- "reference": "ghcr.io/eliovp/paiton-vllm-plugin@sha256:8866c8dd29d7dfbba81156a83af75229ffc9bc05a074833bc516d8da79ebd6a8"
22
  },
23
  "tools": {
24
- "tag": "ghcr.io/eliovp/paiton-vllm-plugin:flux2-tools-rdna4-v1.0.0",
25
- "digest": "sha256:9268cb0505ba4db59adfa7f7faf87ccabf3803571757add7397da40752f453ba",
26
- "reference": "ghcr.io/eliovp/paiton-vllm-plugin@sha256:9268cb0505ba4db59adfa7f7faf87ccabf3803571757add7397da40752f453ba"
27
  },
28
  "comfyui": {
29
- "tag": "ghcr.io/eliovp/paiton-vllm-plugin:flux2-comfyui-rdna4-v1.0.0",
30
- "digest": "sha256:fdede6c2012be50b40207a04a6ed38b79b045b49267cf89180a0577e09de6fff",
31
- "reference": "ghcr.io/eliovp/paiton-vllm-plugin@sha256:fdede6c2012be50b40207a04a6ed38b79b045b49267cf89180a0577e09de6fff"
32
  }
33
  },
34
  "scope": {
 
1
  {
2
  "format_version": 1,
3
  "repository": "EliovpAI/FLUX.2-klein-4B-Paiton-RDNA4",
4
+ "version": "v1.0.1",
5
  "content": "Compiled runtime artifacts, manifests, notices, benchmark data and example media. No checkpoint weights.",
6
  "runtime_source": {
7
  "repository": "https://github.com/Eliovp-BV/paiton-vllm-plugin",
8
+ "revision": "56be0226cb18fec05c3d562b57bcbd752a9c66ea",
9
+ "tag": "paiton-flux2-klein-gfx1201-v1.0.1",
10
  "path": "models/FLUX.2-klein"
11
  },
12
  "source_checkpoint": {
 
16
  },
17
  "containers": {
18
  "klein": {
19
+ "tag": "ghcr.io/eliovp/paiton-vllm-plugin:flux2-klein-rdna4-v1.0.1",
20
+ "digest": "sha256:994af1fdfa53d1f0b0a398bb247d4c1ac6caa93f4a56f22f90d961ca76e1f7de",
21
+ "reference": "ghcr.io/eliovp/paiton-vllm-plugin@sha256:994af1fdfa53d1f0b0a398bb247d4c1ac6caa93f4a56f22f90d961ca76e1f7de"
22
  },
23
  "tools": {
24
+ "tag": "ghcr.io/eliovp/paiton-vllm-plugin:flux2-tools-rdna4-v1.0.1",
25
+ "digest": "sha256:d9ffefd705d5f010c65199a88a0240275aed06f2042a9b1f1d5b5a144d013e61",
26
+ "reference": "ghcr.io/eliovp/paiton-vllm-plugin@sha256:d9ffefd705d5f010c65199a88a0240275aed06f2042a9b1f1d5b5a144d013e61"
27
  },
28
  "comfyui": {
29
+ "tag": "ghcr.io/eliovp/paiton-vllm-plugin:flux2-comfyui-rdna4-v1.0.1",
30
+ "digest": "sha256:b00940f1ee01ce4061b392395458aa0a3a48c51dad69794e16a5f24b8b20add3",
31
+ "reference": "ghcr.io/eliovp/paiton-vllm-plugin@sha256:b00940f1ee01ce4061b392395458aa0a3a48c51dad69794e16a5f24b8b20add3"
32
  }
33
  },
34
  "scope": {