Text Generation
MLX
Safetensors
phi3
4-bit precision
4bit
apple-silicon
chat
conversational
edge-ai
efficient
fast
function-calling
instruct
local-llm
m1
m2
m3
m4
mac
mac-mini
mac-studio
macbook-air
macbook-pro
macos
metal
microsoft
mlx-community
mlx-lm
no-cloud
offline
on-device
outlier
outlier-app
phi
phi-4
phi4
private
private-ai
quantized
small-llm
tool-use
custom_code
Instructions to use Outlier-Ai/Phi-4-mini-instruct-MLX-4bit with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- MLX
How to use Outlier-Ai/Phi-4-mini-instruct-MLX-4bit with MLX:
# Make sure mlx-lm is installed # pip install --upgrade mlx-lm # Generate text with mlx-lm from mlx_lm import load, generate model, tokenizer = load("Outlier-Ai/Phi-4-mini-instruct-MLX-4bit") prompt = "Write a story about Einstein" messages = [{"role": "user", "content": prompt}] prompt = tokenizer.apply_chat_template( messages, add_generation_prompt=True ) text = generate(model, tokenizer, prompt=prompt, verbose=True) - Notebooks
- Google Colab
- Kaggle
- Local Apps Settings
- LM Studio
- MLX LM
How to use Outlier-Ai/Phi-4-mini-instruct-MLX-4bit with MLX LM:
Generate or start a chat session
# Install MLX LM uv tool install mlx-lm # Interactive chat REPL mlx_lm.chat --model "Outlier-Ai/Phi-4-mini-instruct-MLX-4bit"
Run an OpenAI-compatible server
# Install MLX LM uv tool install mlx-lm # Start the server mlx_lm.server --model "Outlier-Ai/Phi-4-mini-instruct-MLX-4bit" # Calling the OpenAI-compatible server with curl curl -X POST "http://localhost:8000/v1/chat/completions" \ -H "Content-Type: application/json" \ --data '{ "model": "Outlier-Ai/Phi-4-mini-instruct-MLX-4bit", "messages": [ {"role": "user", "content": "Hello"} ] }' - Atomic Chat
| license: mit | |
| pipeline_tag: text-generation | |
| library_name: mlx | |
| base_model: microsoft/Phi-4-mini-instruct | |
| base_model_relation: quantized | |
| quantized_by: Outlier-Ai | |
| tags: | |
| - 4-bit | |
| - 4bit | |
| - apple-silicon | |
| - chat | |
| - conversational | |
| - edge-ai | |
| - efficient | |
| - fast | |
| - function-calling | |
| - instruct | |
| - local-llm | |
| - m1 | |
| - m2 | |
| - m3 | |
| - m4 | |
| - mac | |
| - mac-mini | |
| - mac-studio | |
| - macbook-air | |
| - macbook-pro | |
| - macos | |
| - metal | |
| - microsoft | |
| - mlx | |
| - mlx-community | |
| - mlx-lm | |
| - no-cloud | |
| - offline | |
| - on-device | |
| - outlier | |
| - outlier-app | |
| - phi | |
| - phi-4 | |
| - phi4 | |
| - private | |
| - private-ai | |
| - quantized | |
| - safetensors | |
| - small-llm | |
| - text-generation | |
| - tool-use | |
| language: | |
| - en | |
| - de | |
| - fr | |
| - es | |
| - it | |
| - pt | |
| - nl | |
| - pl | |
| widget: | |
| - example_title: Quick answer | |
| messages: | |
| - role: user | |
| content: What is the capital of France? | |
| - example_title: Short definition | |
| messages: | |
| - role: user | |
| content: Define 'mixture of experts' in one sentence. | |
| - example_title: Quick list | |
| messages: | |
| - role: user | |
| content: Name three reasons to run a local LLM. | |
| > **Run this on your Mac with [Outlier](https://outlier.host/?utm_source=hf&utm_medium=modelcard&utm_campaign=phi_4_mini_instruct_mlx_4bit)** β a free macOS app for local MLX inference. | |
| # Phi-4-mini-instruct (MLX 4-bit) | |
| MLX 4-bit conversion of [`microsoft/Phi-4-mini-instruct`](https://huggingface.co/microsoft/Phi-4-mini-instruct). License and base-model fields inherit from the original β see YAML frontmatter above. | |
| ## Load with mlx-lm | |
| ```bash | |
| pip install mlx-lm | |
| python -m mlx_lm.generate --model Outlier-Ai/Phi-4-mini-instruct-MLX-4bit --prompt "Hello" --max-tokens 256 | |
| ``` | |
| ## What is Outlier? | |
| A free macOS app that runs MLX models locally β no cloud, no API keys, no usage caps. | |
| β‘ **[outlier.host](https://outlier.host/?utm_source=hf&utm_medium=modelcard&utm_campaign=phi_4_mini_instruct_mlx_4bit)** | |
| ## Other Outlier conversions | |
| - [DeepSeek-R1-Distill-Qwen-7B (MLX 4-bit) β MLX 4-bit conversion (1,932 downloads)](https://huggingface.co/Outlier-Ai/DeepSeek-R1-Distill-Qwen-7B-MLX-4bit) | |
| - [Qwen3-Coder-30B-A3B-Instruct (MLX 4-bit) β MLX 4-bit conversion (1,598 downloads)](https://huggingface.co/Outlier-Ai/Qwen3-Coder-30B-A3B-Instruct-MLX-4bit) | |
| - [Qwen2.5-Coder-7B-Instruct (MLX 4-bit) β MLX 4-bit conversion (1,450 downloads)](https://huggingface.co/Outlier-Ai/Qwen2.5-Coder-7B-Instruct-MLX-4bit) | |
| - [Outlier-Core-27B (MLX 4-bit) β MLX 4-bit conversion (55 downloads)](https://huggingface.co/Outlier-Ai/Outlier-Core-27B-MLX-4bit) | |
| - [Outlier-Nano-4B (MLX 4-bit) β MLX 4-bit conversion (67 downloads)](https://huggingface.co/Outlier-Ai/Outlier-Nano-4B-MLX-4bit) | |
| ## License | |
| Inherits from upstream (`mit`). See base model card. | |