micdn commited on
Commit
9b4b91a
·
verified ·
1 Parent(s): f0d8ac5

Upload README.md with huggingface_hub

Browse files
Files changed (1) hide show
  1. README.md +133 -0
README.md ADDED
@@ -0,0 +1,133 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ ---
2
+ library_name: mesh-llm
3
+ base_model:
4
+ - "unsloth/inkling-GGUF"
5
+ pipeline_tag: "image-text-to-text"
6
+ tags:
7
+ - gguf
8
+ - mesh-llm
9
+ - layer-package
10
+ - skippy
11
+ - distributed-inference
12
+ - local-inference
13
+ - openai-compatible
14
+ - experimental
15
+ ---
16
+
17
+ <div align="center">
18
+ <a href="https://www.meshllm.cloud">
19
+ <img src="https://meshllm.cloud/assets/images/jelly-logo-wordmark.png" alt="Mesh LLM" width="220">
20
+ </a>
21
+
22
+ <h1>inkling-UD-Q2_K_XL</h1>
23
+
24
+ <p>
25
+ <strong>Distributed GGUF inference package for Mesh LLM</strong>
26
+ </p>
27
+
28
+ <p>
29
+ <a href="https://www.meshllm.cloud"><img alt="Website" src="https://img.shields.io/badge/Website-meshllm.cloud-111111?style=for-the-badge"></a>
30
+ <a href="https://github.com/Mesh-LLM/mesh-llm"><img alt="GitHub" src="https://img.shields.io/badge/GitHub-Mesh--LLM-24292f?style=for-the-badge&logo=github"></a>
31
+ <a href="https://discord.gg/rs6fmc63eN"><img alt="Discord" src="https://img.shields.io/badge/Discord-Join-5865F2?style=for-the-badge&logo=discord&logoColor=white"></a>
32
+ </p>
33
+ </div>
34
+
35
+ > [!WARNING]
36
+ > **Experimental package:** artifact integrity may be validated, but runtime, split-correctness, and multimodal certification are still pending. This package is not discoverable through `meshllm/catalog@main` until its Hugging Face catalog PR is reviewed and merged.
37
+
38
+ GGUF layer package for running **inkling-UD-Q2_K_XL** across a local Mesh LLM cluster.
39
+
40
+ This package is derived from [unsloth/inkling-GGUF](https://huggingface.co/unsloth/inkling-GGUF) and keeps the original GGUF distribution split into per-layer artifacts for distributed inference.
41
+
42
+ ## Highlights
43
+
44
+ | Run locally | Pool multiple machines | OpenAI-compatible | Package variant |
45
+ |---|---|---|---|
46
+ | Private inference on your hardware | Split layers across peers | Serve `/v1/chat/completions` locally | `UD-Q2_K_XL` layer package |
47
+
48
+ ## Model Overview
49
+
50
+ | Property | Value |
51
+ |---|---|
52
+ | **Source model** | [unsloth/inkling-GGUF](https://huggingface.co/unsloth/inkling-GGUF) |
53
+ | **Model id** | `unsloth/inkling-GGUF:UD-Q2_K_XL` |
54
+ | **Family** | inkling |
55
+ | **Parameter scale** | not recorded |
56
+ | **Quantization** | `UD-Q2_K_XL` |
57
+ | **Layer count** | 66 |
58
+ | **Activation width** | 6144 |
59
+ | **Package size** | 296.5 GB |
60
+ | **Source file** | `UD-Q2_K_XL/inkling-UD-Q2_K_XL-00001-of-00008.gguf` |
61
+ | **Package repo** | [meshllm/inkling-UD-Q2_K_XL-layers](https://huggingface.co/meshllm/inkling-UD-Q2_K_XL-layers) |
62
+
63
+ ## Recommended Use
64
+
65
+ - Local and private inference with Mesh LLM.
66
+ - Multi-machine serving when the full GGUF is too large for one host.
67
+ - OpenAI-compatible chat/completions workflows through Mesh LLM's local API.
68
+
69
+ For upstream architecture details, chat template guidance, sampling recommendations, license terms, and benchmark notes, see the source model card: [unsloth/inkling-GGUF](https://huggingface.co/unsloth/inkling-GGUF).
70
+
71
+ ## Quickstart
72
+
73
+ ```bash
74
+ # Run this on each machine that should contribute memory/compute.
75
+ mesh-llm serve --model "meshllm/inkling-UD-Q2_K_XL-layers" --split
76
+ ```
77
+
78
+ ```bash
79
+ # Check the mesh and discover the OpenAI-compatible model name.
80
+ curl -s http://localhost:3131/api/status
81
+ curl -s http://localhost:3131/v1/models
82
+ ```
83
+
84
+ ```bash
85
+ # Send an OpenAI-compatible chat request.
86
+ curl -s http://localhost:3131/v1/chat/completions \
87
+ -H "Content-Type: application/json" \
88
+ -d '{
89
+ "model": "unsloth/inkling-GGUF:UD-Q2_K_XL",
90
+ "messages": [{"role": "user", "content": "Write a tiny hello-world function in Rust."}],
91
+ "max_tokens": 128
92
+ }'
93
+ ```
94
+
95
+ ## Package Variant
96
+
97
+ | Property | Value |
98
+ |---|---|
99
+ | **Format** | `layer-package` |
100
+ | **Canonical source ref** | `unsloth/inkling-GGUF@d3e9ffca48751dbe8b59dab5cfa364621257c682/UD-Q2_K_XL/inkling-UD-Q2_K_XL-00001-of-00008.gguf` |
101
+ | **Source revision** | `d3e9ffca48751dbe8b59dab5cfa364621257c682` |
102
+ | **Source SHA-256** | `8b153eb6a470303f227ef2ba5b515e4e53cc51be65f4395490aa6e37f377bebf` |
103
+ | **Skippy ABI** | `0.1.30` |
104
+ | **Package manifest SHA-256** | `6cda4ce683e8046818d0d51c57e00b61d2245c36ba7b21a4fb7c51ab82c039b7` |
105
+
106
+ ## What Is Included
107
+
108
+ | Artifact | Path | Contents | SHA-256 |
109
+ |---|---|---|---|
110
+ | Manifest | `model-package.json` | Package schema, source identity, checksums | `6cda4ce683e8046818d0d51c57e00b61d2245c36ba7b21a4fb7c51ab82c039b7` |
111
+ | Metadata | `shared/metadata.gguf` | 1 tensors, 12.4 MB | `c2332ae2f8fc2a711861bae5c96720592aee75cdf7868a3bade02bc513f03bf5` |
112
+ | Embeddings | `shared/embeddings.gguf` | 2 tensors, 822.2 MB | `2021ca206c89e66637e4e048caea5c5b181413ce4e9b045fb6aff8db315f7a7a` |
113
+ | Output head | `shared/output.gguf` | 3 tensors, 675.0 MB | `0ca256c69c9c311dfb5671643f9edaff62de8706453a68aa34b953aaac6f8b21` |
114
+ | Transformer layers | `layers/layer-*.gguf` | 66 layer artifacts, 1574 tensors, 294.9 GB | `see model-package.json` |
115
+ | Projector | `projectors/mmproj-BF16.gguf` | mmproj projector, 174.8 MB | `662c925e1df293cfba16ffd6bd53dac31d3c73160ba65dff7270d7a70f351e91` |
116
+
117
+ ## Validation
118
+
119
+ Generated by the Mesh LLM HF Jobs splitter from `mesh-llm` ref `codex/inkling-q2-skippy`.
120
+ Each artifact is checksummed as it is written, uploaded to this repository, and removed from the job workspace before the next artifact is produced.
121
+
122
+ ```bash
123
+ skippy-model-package write-package "/source/UD-Q2_K_XL/inkling-UD-Q2_K_XL-00001-of-00008.gguf" --out-dir "/tmp/meshllm-layer-job-meshllm_inkling-UD-Q2_K_XL-layers-193/package"
124
+ ```
125
+
126
+ ## Links
127
+
128
+ - Source model: [unsloth/inkling-GGUF](https://huggingface.co/unsloth/inkling-GGUF)
129
+ - Mesh LLM website: [meshllm.cloud](https://www.meshllm.cloud)
130
+ - Mesh LLM: [github.com/Mesh-LLM/mesh-llm](https://github.com/Mesh-LLM/mesh-llm)
131
+ - Discord: [discord.gg/rs6fmc63eN](https://discord.gg/rs6fmc63eN)
132
+ - Package catalog: [meshllm/catalog](https://huggingface.co/datasets/meshllm/catalog)
133
+ - Package format: [layer-package-repos.md](https://github.com/Mesh-LLM/mesh-llm/blob/main/docs/specs/layer-package-repos.md)