Instructions to use tatjr13/damascus-t3-base-maxprec-shipped with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Notebooks
- Google Colab
- Kaggle
- Local Apps Settings
- llama.cpp
How to use tatjr13/damascus-t3-base-maxprec-shipped with llama.cpp:
Install (macOS, Linux)
curl -LsSf https://llama.app/install.sh | sh # Start a local OpenAI-compatible server with a web UI: llama serve -hf tatjr13/damascus-t3-base-maxprec-shipped # Run inference directly in the terminal: llama cli -hf tatjr13/damascus-t3-base-maxprec-shipped
Install from WinGet (Windows)
winget install llama.cpp # Start a local OpenAI-compatible server with a web UI: llama serve -hf tatjr13/damascus-t3-base-maxprec-shipped # Run inference directly in the terminal: llama cli -hf tatjr13/damascus-t3-base-maxprec-shipped
Use pre-built binary
# Download pre-built binary from: # https://github.com/ggerganov/llama.cpp/releases # Start a local OpenAI-compatible server with a web UI: ./llama-server -hf tatjr13/damascus-t3-base-maxprec-shipped # Run inference directly in the terminal: ./llama-cli -hf tatjr13/damascus-t3-base-maxprec-shipped
Build from source code
git clone https://github.com/ggerganov/llama.cpp.git cd llama.cpp cmake -B build cmake --build build -j --target llama-server llama-cli # Start a local OpenAI-compatible server with a web UI: ./build/bin/llama-server -hf tatjr13/damascus-t3-base-maxprec-shipped # Run inference directly in the terminal: ./build/bin/llama-cli -hf tatjr13/damascus-t3-base-maxprec-shipped
Use Docker
docker model run hf.co/tatjr13/damascus-t3-base-maxprec-shipped
- LM Studio
- Jan
- Ollama
How to use tatjr13/damascus-t3-base-maxprec-shipped with Ollama:
ollama run hf.co/tatjr13/damascus-t3-base-maxprec-shipped
- Unsloth Desktop
- Docker Model Runner
How to use tatjr13/damascus-t3-base-maxprec-shipped with Docker Model Runner:
docker model run hf.co/tatjr13/damascus-t3-base-maxprec-shipped
- Lemonade
How to use tatjr13/damascus-t3-base-maxprec-shipped with Lemonade:
Pull the model
# Download Lemonade from https://lemonade-server.ai/ lemonade pull tatjr13/damascus-t3-base-maxprec-shipped
Run and chat with the model
lemonade run user.damascus-t3-base-maxprec-shipped-{{QUANT_TAG}}List all available models
lemonade list
- Atomic Chat
YAML Metadata Warning:empty or missing yaml metadata in repo card
Check out the documentation for more information.
damascus-t3-base-maxprec-shipped
Untouched tpn-004 base (Magistral-Small-2509-ultra-uncensored-heretic-v2, tpnlabs/tpn-004-base BF16, sha 5d425b36...) exported with the DAMASCUS T3 "maximum counted precision under the cap" layout (TPN-200): the L1 recipe re-spent so every validator-counted byte goes where KL divergence says the model is most sensitive โ early30 Q6_K only (late stays Q4_K_M); funded by token_embd Q8_0->Q6_K, ffn_down Q8_0->Q5_K โ with the SHIPPED 305-char stripped chat template used by tatjr13/tpn004-cand-v4f-l1im (template sha 885e0c70...). Quantized from the BF16 with the evaluator-shaped-v1 imatrix via llama.cpp b10020 llama-quantize.
Measured on the b6 KL corpus (150 epoch-p1 MCQ items through the shipped template + 150 general chunks, sha f30e711c..., 96,145 tokens, ctx 512, all KL passes on llama.cpp b10453/Vulkan): mean KL vs the BF16 base 0.003088 (L1 recipe anchor: 0.006227; all local, UNCALIBRATED). Validator-counted memory (b10020 CPU probe, TPN_PROBE_CAP_KB=16777216): 16,636,497 KiB at ctx 4096 (3/3 identical, cap pass) and 17,291,857 KiB at ctx 8192 (over the same cap at 8192; T3's constraint is the 4096 figure). Test artifact, not a competition entry.
- Downloads last month
- 137
We're not able to determine the quantization variants.