Instructions to use mykor/harrier-oss-v1-270m-GGUF with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Notebooks
- Google Colab
- Kaggle
- Local Apps Settings
- llama.cpp
How to use mykor/harrier-oss-v1-270m-GGUF with llama.cpp:
Install (macOS, Linux)
curl -LsSf https://llama.app/install.sh | sh # Start a local OpenAI-compatible server with a web UI: llama serve -hf mykor/harrier-oss-v1-270m-GGUF:Q4_K_M # Run inference directly in the terminal: llama cli -hf mykor/harrier-oss-v1-270m-GGUF:Q4_K_M
Install from WinGet (Windows)
winget install llama.cpp # Start a local OpenAI-compatible server with a web UI: llama serve -hf mykor/harrier-oss-v1-270m-GGUF:Q4_K_M # Run inference directly in the terminal: llama cli -hf mykor/harrier-oss-v1-270m-GGUF:Q4_K_M
Use pre-built binary
# Download pre-built binary from: # https://github.com/ggerganov/llama.cpp/releases # Start a local OpenAI-compatible server with a web UI: ./llama-server -hf mykor/harrier-oss-v1-270m-GGUF:Q4_K_M # Run inference directly in the terminal: ./llama-cli -hf mykor/harrier-oss-v1-270m-GGUF:Q4_K_M
Build from source code
git clone https://github.com/ggerganov/llama.cpp.git cd llama.cpp cmake -B build cmake --build build -j --target llama-server llama-cli # Start a local OpenAI-compatible server with a web UI: ./build/bin/llama-server -hf mykor/harrier-oss-v1-270m-GGUF:Q4_K_M # Run inference directly in the terminal: ./build/bin/llama-cli -hf mykor/harrier-oss-v1-270m-GGUF:Q4_K_M
Use Docker
docker model run hf.co/mykor/harrier-oss-v1-270m-GGUF:Q4_K_M
- LM Studio
- Jan
- Ollama
How to use mykor/harrier-oss-v1-270m-GGUF with Ollama:
ollama run hf.co/mykor/harrier-oss-v1-270m-GGUF:Q4_K_M
- Unsloth Desktop
- Docker Model Runner
How to use mykor/harrier-oss-v1-270m-GGUF with Docker Model Runner:
docker model run hf.co/mykor/harrier-oss-v1-270m-GGUF:Q4_K_M
- Lemonade
How to use mykor/harrier-oss-v1-270m-GGUF with Lemonade:
Pull the model
# Download Lemonade from https://lemonade-server.ai/ lemonade pull mykor/harrier-oss-v1-270m-GGUF:Q4_K_M
Run and chat with the model
lemonade run user.harrier-oss-v1-270m-GGUF-Q4_K_M
List all available models
lemonade list
- Atomic Chat
harrier-oss-v1-270m-GGUF
import numpy as np
from llama_cpp import Llama
from sentence_transformers import SentenceTransformer
from sentence_transformers.util import cos_sim
model = SentenceTransformer(
"microsoft/harrier-oss-v1-270m",
)
llama = Llama.from_pretrained(
repo_id="mykor/harrier-oss-v1-270m-GGUF",
filename="harrier-oss-v1-270M-BF16.gguf",
verbose=False,
embedding=True,
n_ctx=0,
)
text = """๊ธฐํ ์ค์ด ๊ดํ ํผ์ ์ธ๋ฆฌ๋ ์๋ฆฌ๊ฐ
๋ ์ผ์ผํค๋ ๊ฒ ๊ฐ์
์ค๋๋ฐ๋ผ ๋ถ์ ๊บผ์ ธ ํ์ด์ ๋ฐ๋
๋ฐฉ์์ด ๋ญ๊ฐ ํ์ ํด
๋ฐ๋์ ๊ธฐ์ตํด ์ค๋
ํฅ์ผ๊ฑฐ๋ฆด ๋ ์ด ๋ง์ ๋ฐ๋ผ๋๊น
์กฐ๊ธ ๋ ๋ฉ๋ฆฌ ๋ฉ๋ฆฌ์ ์ด๋๋ ๋ฟ๊ฒ
๋ด ์ ๋ถ๋ฅผ ์ ๋ถ ๋ค ์ค๊ฒ
์๋ฌด๊ฒ๋ ๋ค๋ฆฌ์ง ์๋๋ผ๋
๋ด ์์ ๋๋ฅผ ์ํ ์์ด ๋ค๋ ค
์ค๋ ์ ๋ ๋ฏธ๋ฃจ๊ธฐ ์ซ๋คํด๋
์ฐ์ ๋ง์ดํฌ์๋ง ์์ญ์ผ๊ฒ
Ooh-oh, ooh-oh
์ข ๋ฏธ์ํด ์์ง ์ค๋น๊ฐ ์ ๋ ๊ฒ ๊ฐ์
๊ฐ๋์ ํผ์ ์ฌ๊ณค ํ์ด
๋์ด๋
ผ ๋ง๋ค์ ์ด๋ฏธ ์ ๋ถ ๋ง๋ผ
๋ด๊ฐ ๊ฐ ๊ณณ์ ์ ํ์ผ๋๊น
๋ฌ๋น์ ๊ธฐ์ตํด์ค๋
๊ฟ์์ ๋ชฐ๋ ๋ถ๋ฅผ์ง ๋ชจ๋ฅด๋๊น
์กฐ๊ธ ๋ ๋ฉ๋ฆฌ ๋ฉ๋ฆฌ์ ์ด๋๋ ๋ฟ๊ฒ
๋ด ์ ๋ถ๋ฅผ ์ ๋ถ ๋ค ์ค๊ฒ
์๋ฌด๊ฒ๋ ๋ค๋ฆฌ์ง ์๋๋ผ๋
๋ด ์์ ๋๋ฅผ ์ํ ์์ด ๋ค๋ ค
์ค๋ ์ ๋ ๋ฏธ๋ฃจ๊ธฐ ์ซ๋คํด๋
์ฐ์ ๋ง์ดํฌ์๋ง ์์ญ์ผ๊ฒ
Ooh-oh (oh) ooh-oh (oh) ooh-oh"""
embed1 = model.encode(text)
embed2 = np.array(llama.embed(text), dtype=np.float32)
print(cos_sim(embed1, embed2).item())
0.9999479055404663
- Downloads last month
- 1,401
Hardware compatibility
Log In to add your hardware
3-bit
4-bit
5-bit
6-bit
8-bit
16-bit
Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐ Ask for provider support
Model tree for mykor/harrier-oss-v1-270m-GGUF
Base model
microsoft/harrier-oss-v1-270m