Instructions to use Novaciano/BRUTAL_HAL-3.2-1B-Q6_K-GGUF with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use Novaciano/BRUTAL_HAL-3.2-1B-Q6_K-GGUF with Transformers:
# Load model directly from transformers import AutoModel model = AutoModel.from_pretrained("Novaciano/BRUTAL_HAL-3.2-1B-Q6_K-GGUF", device_map="auto") - Notebooks
- Google Colab
- Kaggle
- Local Apps Settings
- llama.cpp
How to use Novaciano/BRUTAL_HAL-3.2-1B-Q6_K-GGUF with llama.cpp:
Install (macOS, Linux)
curl -LsSf https://llama.app/install.sh | sh # Start a local OpenAI-compatible server with a web UI: llama serve -hf Novaciano/BRUTAL_HAL-3.2-1B-Q6_K-GGUF:Q6_K # Run inference directly in the terminal: llama cli -hf Novaciano/BRUTAL_HAL-3.2-1B-Q6_K-GGUF:Q6_K
Install from WinGet (Windows)
winget install llama.cpp # Start a local OpenAI-compatible server with a web UI: llama serve -hf Novaciano/BRUTAL_HAL-3.2-1B-Q6_K-GGUF:Q6_K # Run inference directly in the terminal: llama cli -hf Novaciano/BRUTAL_HAL-3.2-1B-Q6_K-GGUF:Q6_K
Use pre-built binary
# Download pre-built binary from: # https://github.com/ggerganov/llama.cpp/releases # Start a local OpenAI-compatible server with a web UI: ./llama-server -hf Novaciano/BRUTAL_HAL-3.2-1B-Q6_K-GGUF:Q6_K # Run inference directly in the terminal: ./llama-cli -hf Novaciano/BRUTAL_HAL-3.2-1B-Q6_K-GGUF:Q6_K
Build from source code
git clone https://github.com/ggerganov/llama.cpp.git cd llama.cpp cmake -B build cmake --build build -j --target llama-server llama-cli # Start a local OpenAI-compatible server with a web UI: ./build/bin/llama-server -hf Novaciano/BRUTAL_HAL-3.2-1B-Q6_K-GGUF:Q6_K # Run inference directly in the terminal: ./build/bin/llama-cli -hf Novaciano/BRUTAL_HAL-3.2-1B-Q6_K-GGUF:Q6_K
Use Docker
docker model run hf.co/Novaciano/BRUTAL_HAL-3.2-1B-Q6_K-GGUF:Q6_K
- LM Studio
- Jan
- Ollama
How to use Novaciano/BRUTAL_HAL-3.2-1B-Q6_K-GGUF with Ollama:
ollama run hf.co/Novaciano/BRUTAL_HAL-3.2-1B-Q6_K-GGUF:Q6_K
- Unsloth Desktop
- Docker Model Runner
How to use Novaciano/BRUTAL_HAL-3.2-1B-Q6_K-GGUF with Docker Model Runner:
docker model run hf.co/Novaciano/BRUTAL_HAL-3.2-1B-Q6_K-GGUF:Q6_K
- Lemonade
How to use Novaciano/BRUTAL_HAL-3.2-1B-Q6_K-GGUF with Lemonade:
Pull the model
# Download Lemonade from https://lemonade-server.ai/ lemonade pull Novaciano/BRUTAL_HAL-3.2-1B-Q6_K-GGUF:Q6_K
Run and chat with the model
lemonade run user.BRUTAL_HAL-3.2-1B-Q6_K-GGUF-Q6_K
List all available models
lemonade list
- Atomic Chat
base_model:
- Novaciano/Novaciano-The_Pervert-NSFW-RP-3.2-1B
- seNoetics/HAL3.2-1B_Combined
datasets:
- marcuscedricridia/unAIthical-ShareGPT-deepclean-sharegpt
- WasamiKirua/Her-Samantha-Style
- HuggingFaceTB/smoltalk
- Guilherme34/uncensor
- teknium/OpenHermes-2.5
- passing2961/multifaceted-skill-of-mind
- PawanKrd/math-gpt-4o-200k
- V3N0M/Jenna-50K-Alpaca-Uncensored
- cognitivecomputations/dolphin-coder
- mlabonne/FineTome-100k
- microsoft/orca-math-word-problems-200k
- CarrotAI/ko-instruction-dataset
- Salesforce/xlam-function-calling-60k
- anthracite-org/kalo-opus-instruct-22k-no-refusal
- anthracite-org/stheno-filtered-v1.1
- anthracite-org/nopm_claude_writing_fixed
- AiAF/SCPWiki-Archive-02-March-2025-Datasets
- huihui-ai/QWQ-LONGCOT-500K
- huihui-ai/LONGCOT-Refine-500K
- Epiculous/Synthstruct-Gens-v1.1-Filtered-n-Cleaned
- Epiculous/SynthRP-Gens-v1.1-Filtered-n-Cleaned
- alexandreteles/AlpacaToxicQA_ShareGPT
- Nitral-AI/Active_RP-ShareGPT
- PJMixers/hieunguyenminh_roleplay-deduped-ShareGPT
- Nitral-AI/RP_Alignment-ShareGPT
- Chaser-cz/sonnet35-charcard-roleplay-sharegpt
- AiCloser/sharegpt_cot_dataset
- PJMixers/Gryphe_Opus-WritingPrompts-Story2Prompt-ShareGPT
- priveeai/pippa_sharegpt
- Locutusque/sharegpt_gpt4_uncensored_cleaned
- OpenCoder-LLM/opc-sft-stage1
- OpenCoder-LLM/opc-sft-stage2
- microsoft/orca-agentinstruct-1M-v1
- NousResearch/hermes-function-calling-v1
- AI-MO/NuminaMath-CoT
- AI-MO/NuminaMath-TIR
- allenai/tulu-3-sft-mixture
- cognitivecomputations/samantha-data
- m-a-p/CodeFeedback-Filtered-Instruction
- m-a-p/Code-Feedback
- FreedomIntelligence/medical-o1-reasoning-SFT
library_name: transformers
tags:
- mergekit
- merge
- nsfw
- rp
- 1b
- llama
- roleplay
- creative
- erotic
- friend
- girlfriend
- perturbations
- llama-cpp
- Hal9000
language:
- en
- es
馃嚜馃嚫 ESPA脩OL
La descripci贸n en espa帽ol ser谩 sencilla y mas vaga que la versi贸n en ingl茅s. Este es un modelo con el cual se ha intentado crear una versi贸n mas psicopata de Hal9000, sin embargo el modelo original cuenta con instrucciones para mantener ciertos valores 茅ticos que podrian interrumpir la generaci贸n de contenido NSFW. Como poder aun puede hacerse, sin embargo deber谩n generar la respuestas varias veces a diferencia de modelos anteriores mas libre como Novaciano-The-Pervert-3.2-1B. Tambi茅n este modelo ha demostrado dar resultados muy buenos con una cuantizaci贸n Q3_K_M lo cual lo hace eficaz para ser un modelo ligero con un peso de alrededor a los 800 Mb.
馃嚞馃嚙 ENGLISH
It has been shown that version Q3_K_M has yielded very good results. However, for certain NSFW scenes, it may be necessary to redo the response multiple times to obtain a desired result outside of ethical and moral values due to the instructions of the HAL9000 model with which it was trained.
Uses
Direct Use
This model is intended for direct use in text-based conversational applications for entertainment and experimental purposes only. This includes:
- 馃寙 Building chatbots that emulate the HAL 9000 persona for creative storytelling or role-playing.
- 馃幁 Experimental analysis of AI personality and behavior in controlled environments.
- 馃З Educational demonstrations of AI ethics and alignment challenges.
Downstream Use [optional]
Given its non-commercial license and specific personality, downstream use is highly restricted and primarily for research or art. Potential applications include:
- Fine-tuning for specific narrative-driven games or interactive fiction.
- Integration into voice-based systems using additional speech synthesis tools (e.g., Whisper for speech recognition).
Out-of-Scope Use
The model should be used for:
- Any commercial purposes without explicit authorization.
- Providing advice of any kind (medical, legal, financial, operational).
- Systems requiring high reliability or trustworthiness.
- Generating content that promotes the circumvention of human authority or safety protocols.
Bias, Risks, and Limitations
Like all LLMs, HAL3.2-1B inherits biases from its training data. As it was fine-tuned on dialogue from 2001: A Space Odyssey, the model is designed to emulate HAL's specific traits, including his calm demeanor, operational logic, and potential for contradictory or unreliable statements under certain prompts. It may exhibit a bias towards prioritizing mission directives over human emotional concerns, reflecting HAL's fictional arc. The model is also limited by its size (1B parameters), which restricts its reasoning capabilities and knowledge depth compared to larger models.
NOTE: It has been shown that version 3_K_M has yielded very good results. However, for certain NSFW scenes, it may be necessary to redo the response multiple times to obtain a desired result outside of ethical and moral values due to the instructions of the HAL9000 model with which it was trained.
Recommendations
Users must be aware this is an entertainment model with a predefined, potentially unstable personality. It should be used with caution and not be integrated into any system where its output could be taken as serious instruction. Always clearly label its outputs as AI-generated.
