ProCreations's picture
Add verified lossless shards of official MiniCPM5-2B Q8_0 for WebGPU
0f58fbf verified
|
Raw History Blame Contribute Delete
1.11 kB
---
license: apache-2.0
base_model: openbmb/MiniCPM5-2B-GGUF
library_name: gguf
pipeline_tag: text-generation
tags:
- webgpu
- wllama
- minicpm5
- gguf
- q8_0
---
# MiniCPM5-2B · Q8_0 browser shards
A **lossless GGUF split** of OpenBMB's official [MiniCPM5-2B-Q8_0.gguf](https://huggingface.co/openbmb/MiniCPM5-2B-GGUF).
No re-quantization, pruning, conversion of tensor values, or changes to the model architecture.
All 381 tensors were compared by SHA-256 against the original. See `verification.json`.
Split with `llama-gguf-split --split-max-size 512M` into six files, each under the browser's 2 GB ArrayBuffer limit.
Pass the first shard URL to Wllama 3.6.1; it discovers and downloads the other five in parallel.
**Chat and create files entirely in your browser:** [MiniCPM5 WebGPU](https://huggingface.co/spaces/ProCreations/minicpm5-2b-webgpu)
Model by [OpenBMB](https://huggingface.co/openbmb). Browser experience by [ProCreations](https://huggingface.co/ProCreations).
Original weights are Apache 2.0. Please refer to the upstream model card for model capabilities, training, and limitations.