How to use from
Pi
Start the llama.cpp server
# Install llama.cpp:
brew install llama.cpp
# Start a local OpenAI-compatible server:
llama serve -hf TheDrummer/UnslopNemo-12B-v3-GGUF:
Configure the model in Pi
# Install Pi:
npm install -g @earendil-works/pi-coding-agent
# Add to ~/.pi/agent/models.json:
{
  "providers": {
    "llama-cpp": {
      "baseUrl": "http://localhost:8080/v1",
      "api": "openai-completions",
      "apiKey": "none",
      "models": [
        {
          "id": "TheDrummer/UnslopNemo-12B-v3-GGUF:"
        }
      ]
    }
  }
}
Run Pi
# Start Pi in your project directory:
pi
Quick Links

YAML Metadata Warning:empty or missing yaml metadata in repo card

Check out the documentation for more information.

(Previously BeaverAI/Rocinante-12B-v2g-GGUF)

UnslopNemo v3 (Experiment)

I unslopped roughly 90% of my RP dataset in an attempt to make the model more expressive.

Feedback

Usage

  • Metharme (Pygmalion in ST), Mistral, or Text Completion
  • Play around with your samplers. You might have better results if you disable the usual enabled stuff.
Downloads last month
5,350
GGUF
Model size
12B params
Architecture
llama
Hardware compatibility
Log In to add your hardware

2-bit

3-bit

4-bit

5-bit

6-bit

8-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support