Instructions to use tensorblock/RWKV_v6-Finch-14B-HF-GGUF with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Notebooks
- Google Colab
- Kaggle
- Local Apps Settings
- llama.cpp
How to use tensorblock/RWKV_v6-Finch-14B-HF-GGUF with llama.cpp:
Install (macOS, Linux)
curl -LsSf https://llama.app/install.sh | sh # Start a local OpenAI-compatible server with a web UI: llama serve -hf tensorblock/RWKV_v6-Finch-14B-HF-GGUF:Q2_K # Run inference directly in the terminal: llama cli -hf tensorblock/RWKV_v6-Finch-14B-HF-GGUF:Q2_K
Install from WinGet (Windows)
winget install llama.cpp # Start a local OpenAI-compatible server with a web UI: llama serve -hf tensorblock/RWKV_v6-Finch-14B-HF-GGUF:Q2_K # Run inference directly in the terminal: llama cli -hf tensorblock/RWKV_v6-Finch-14B-HF-GGUF:Q2_K
Use pre-built binary
# Download pre-built binary from: # https://github.com/ggerganov/llama.cpp/releases # Start a local OpenAI-compatible server with a web UI: ./llama-server -hf tensorblock/RWKV_v6-Finch-14B-HF-GGUF:Q2_K # Run inference directly in the terminal: ./llama-cli -hf tensorblock/RWKV_v6-Finch-14B-HF-GGUF:Q2_K
Build from source code
git clone https://github.com/ggerganov/llama.cpp.git cd llama.cpp cmake -B build cmake --build build -j --target llama-server llama-cli # Start a local OpenAI-compatible server with a web UI: ./build/bin/llama-server -hf tensorblock/RWKV_v6-Finch-14B-HF-GGUF:Q2_K # Run inference directly in the terminal: ./build/bin/llama-cli -hf tensorblock/RWKV_v6-Finch-14B-HF-GGUF:Q2_K
Use Docker
docker model run hf.co/tensorblock/RWKV_v6-Finch-14B-HF-GGUF:Q2_K
- LM Studio
- Jan
- Ollama
How to use tensorblock/RWKV_v6-Finch-14B-HF-GGUF with Ollama:
ollama run hf.co/tensorblock/RWKV_v6-Finch-14B-HF-GGUF:Q2_K
- Unsloth Desktop
- Docker Model Runner
How to use tensorblock/RWKV_v6-Finch-14B-HF-GGUF with Docker Model Runner:
docker model run hf.co/tensorblock/RWKV_v6-Finch-14B-HF-GGUF:Q2_K
- Lemonade
How to use tensorblock/RWKV_v6-Finch-14B-HF-GGUF with Lemonade:
Pull the model
# Download Lemonade from https://lemonade-server.ai/ lemonade pull tensorblock/RWKV_v6-Finch-14B-HF-GGUF:Q2_K
Run and chat with the model
lemonade run user.RWKV_v6-Finch-14B-HF-GGUF-Q2_K
List all available models
lemonade list
- Atomic Chat
Remove .gguf files (keep Q2_K.gguf)
Browse files- v6-Finch-14B-HF-Q3_K_L.gguf +0 -3
- v6-Finch-14B-HF-Q3_K_M.gguf +0 -3
- v6-Finch-14B-HF-Q3_K_S.gguf +0 -3
- v6-Finch-14B-HF-Q4_0.gguf +0 -3
- v6-Finch-14B-HF-Q4_K_M.gguf +0 -3
- v6-Finch-14B-HF-Q4_K_S.gguf +0 -3
- v6-Finch-14B-HF-Q5_0.gguf +0 -3
- v6-Finch-14B-HF-Q5_K_M.gguf +0 -3
- v6-Finch-14B-HF-Q5_K_S.gguf +0 -3
- v6-Finch-14B-HF-Q6_K.gguf +0 -3
- v6-Finch-14B-HF-Q8_0.gguf +0 -3
v6-Finch-14B-HF-Q3_K_L.gguf
DELETED
|
@@ -1,3 +0,0 @@
|
|
| 1 |
-
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:c9925b519aa5c449a02367f80fb5db4941897b85519f951750114fb2606f1faa
|
| 3 |
-
size 6965251680
|
|
|
|
|
|
|
|
|
|
|
|
v6-Finch-14B-HF-Q3_K_M.gguf
DELETED
|
@@ -1,3 +0,0 @@
|
|
| 1 |
-
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:31fab85bdd0409f9cd83cac75734b2df51490c3356f7dcb4b243026cf4133ba1
|
| 3 |
-
size 6965251680
|
|
|
|
|
|
|
|
|
|
|
|
v6-Finch-14B-HF-Q3_K_S.gguf
DELETED
|
@@ -1,3 +0,0 @@
|
|
| 1 |
-
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:5292082390c49d76cbed8256121d958cbbdaf4046ca4680e90f82c11b5ee13c3
|
| 3 |
-
size 6965251680
|
|
|
|
|
|
|
|
|
|
|
|
v6-Finch-14B-HF-Q4_0.gguf
DELETED
|
@@ -1,3 +0,0 @@
|
|
| 1 |
-
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:8f26b35cc096f1a6f6ad69e511a53a0beb0b68a5063eb9e02d86eb4f8a2f89ee
|
| 3 |
-
size 8767884896
|
|
|
|
|
|
|
|
|
|
|
|
v6-Finch-14B-HF-Q4_K_M.gguf
DELETED
|
@@ -1,3 +0,0 @@
|
|
| 1 |
-
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:e6fdcb718d439b8c8f48e5a3fd21974732875a8c094a99bd51b0151e22bfe68c
|
| 3 |
-
size 8767884896
|
|
|
|
|
|
|
|
|
|
|
|
v6-Finch-14B-HF-Q4_K_S.gguf
DELETED
|
@@ -1,3 +0,0 @@
|
|
| 1 |
-
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:953d42a91eaf2805324198a0d99fc1ffb9714a7214d37ddf750a001de5829a9e
|
| 3 |
-
size 8767884896
|
|
|
|
|
|
|
|
|
|
|
|
v6-Finch-14B-HF-Q5_0.gguf
DELETED
|
@@ -1,3 +0,0 @@
|
|
| 1 |
-
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:21781985a5de3edc0040f2be335d64fa2f8d8a4f9b1a816c446d157839c48899
|
| 3 |
-
size 10464480864
|
|
|
|
|
|
|
|
|
|
|
|
v6-Finch-14B-HF-Q5_K_M.gguf
DELETED
|
@@ -1,3 +0,0 @@
|
|
| 1 |
-
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:20795b7bb4ba924f0b7207889db7c8423fabd1ddffdf62aaaa9bed0b9c214683
|
| 3 |
-
size 10464480864
|
|
|
|
|
|
|
|
|
|
|
|
v6-Finch-14B-HF-Q5_K_S.gguf
DELETED
|
@@ -1,3 +0,0 @@
|
|
| 1 |
-
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:9d6ce78acc8585825e1cacb1ed4284021b5d0d9c8afa83e1d36d1a2017fd450c
|
| 3 |
-
size 10464480864
|
|
|
|
|
|
|
|
|
|
|
|
v6-Finch-14B-HF-Q6_K.gguf
DELETED
|
@@ -1,3 +0,0 @@
|
|
| 1 |
-
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:fd3aad78e7e0e346178ad06b854302cdafbbc86bcee347335a9e31cfc3f7ecf8
|
| 3 |
-
size 12267114080
|
|
|
|
|
|
|
|
|
|
|
|
v6-Finch-14B-HF-Q8_0.gguf
DELETED
|
@@ -1,3 +0,0 @@
|
|
| 1 |
-
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:9405ba24439e95ebe8307590904f56460a3ad8130c3ebd7d38eca286338357e6
|
| 3 |
-
size 15619280480
|
|
|
|
|
|
|
|
|
|
|
|