Instructions to use AutomatosX/AX-Nemotron-3.5-Lightning-30B-A3B-MLX-AXQ-MXFP4-MTP with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- MLX
How to use AutomatosX/AX-Nemotron-3.5-Lightning-30B-A3B-MLX-AXQ-MXFP4-MTP with MLX:
# Make sure mlx-lm is installed # pip install --upgrade mlx-lm # Generate text with mlx-lm from mlx_lm import load, generate model, tokenizer = load("AutomatosX/AX-Nemotron-3.5-Lightning-30B-A3B-MLX-AXQ-MXFP4-MTP") prompt = "Write a story about Einstein" messages = [{"role": "user", "content": prompt}] prompt = tokenizer.apply_chat_template( messages, add_generation_prompt=True ) text = generate(model, tokenizer, prompt=prompt, verbose=True) - Notebooks
- Google Colab
- Kaggle
- Local Apps Settings
- LM Studio
- Pi
How to use AutomatosX/AX-Nemotron-3.5-Lightning-30B-A3B-MLX-AXQ-MXFP4-MTP with Pi:
Start the MLX server
# Install MLX LM: uv tool install mlx-lm # Start a local OpenAI-compatible server: mlx_lm.server --model "AutomatosX/AX-Nemotron-3.5-Lightning-30B-A3B-MLX-AXQ-MXFP4-MTP"
Configure the model in Pi
# Install Pi: npm install -g @earendil-works/pi-coding-agent # Add to ~/.pi/agent/models.json: { "providers": { "mlx-lm": { "baseUrl": "http://localhost:8080/v1", "api": "openai-completions", "apiKey": "none", "models": [ { "id": "AutomatosX/AX-Nemotron-3.5-Lightning-30B-A3B-MLX-AXQ-MXFP4-MTP" } ] } } }Run Pi
# Start Pi in your project directory: pi
- MLX LM
How to use AutomatosX/AX-Nemotron-3.5-Lightning-30B-A3B-MLX-AXQ-MXFP4-MTP with MLX LM:
Generate or start a chat session
# Install MLX LM uv tool install mlx-lm # Interactive chat REPL mlx_lm.chat --model "AutomatosX/AX-Nemotron-3.5-Lightning-30B-A3B-MLX-AXQ-MXFP4-MTP"
Run an OpenAI-compatible server
# Install MLX LM uv tool install mlx-lm # Start the server mlx_lm.server --model "AutomatosX/AX-Nemotron-3.5-Lightning-30B-A3B-MLX-AXQ-MXFP4-MTP" # Calling the OpenAI-compatible server with curl curl -X POST "http://localhost:8000/v1/chat/completions" \ -H "Content-Type: application/json" \ --data '{ "model": "AutomatosX/AX-Nemotron-3.5-Lightning-30B-A3B-MLX-AXQ-MXFP4-MTP", "messages": [ {"role": "user", "content": "Hello"} ] }' - Hermes Agent
How to use AutomatosX/AX-Nemotron-3.5-Lightning-30B-A3B-MLX-AXQ-MXFP4-MTP with Hermes Agent:
Start the MLX server
# Install MLX LM: uv tool install mlx-lm # Start a local OpenAI-compatible server: mlx_lm.server --model "AutomatosX/AX-Nemotron-3.5-Lightning-30B-A3B-MLX-AXQ-MXFP4-MTP"
Configure Hermes
# Install Hermes: curl -fsSL https://hermes-agent.nousresearch.com/install.sh | bash hermes setup # Point Hermes at the local server: hermes config set model.provider custom hermes config set model.base_url http://127.0.0.1:8080/v1 hermes config set model.default AutomatosX/AX-Nemotron-3.5-Lightning-30B-A3B-MLX-AXQ-MXFP4-MTP
Run Hermes
hermes
- Atomic Chat
- OpenClaw
How to use AutomatosX/AX-Nemotron-3.5-Lightning-30B-A3B-MLX-AXQ-MXFP4-MTP with OpenClaw:
Start the MLX server
# Install MLX LM: uv tool install mlx-lm # Start a local OpenAI-compatible server: mlx_lm.server --model "AutomatosX/AX-Nemotron-3.5-Lightning-30B-A3B-MLX-AXQ-MXFP4-MTP"
Configure OpenClaw
# Install OpenClaw: npm install -g openclaw@latest # Register the local server and set it as the default model: openclaw onboard --non-interactive --mode local \ --auth-choice custom-api-key \ --custom-base-url http://127.0.0.1:8080/v1 \ --custom-model-id "AutomatosX/AX-Nemotron-3.5-Lightning-30B-A3B-MLX-AXQ-MXFP4-MTP" \ --custom-provider-id mlx-lm \ --custom-compatibility openai \ --custom-text-input \ --accept-risk \ --skip-health
Run OpenClaw
openclaw agent --local --agent main --message "Hello from Hugging Face"
Download runtime_audit.json from AutomatosX/AX-Nemotron-3.5-Lightning-30B-A3B-MLX-AXQ-MXFP4-MTP: direct link, hf CLI and curl.
- Browser
- Download file 2.89 kB
-
https://huggingface.co/AutomatosX/AX-Nemotron-3.5-Lightning-30B-A3B-MLX-AXQ-MXFP4-MTP/resolve/main/runtime_audit.json
- Command line
-
hf download hf://AutomatosX/AX-Nemotron-3.5-Lightning-30B-A3B-MLX-AXQ-MXFP4-MTP/runtime_audit.json
-
curl -L -o runtime_audit.json https://huggingface.co/AutomatosX/AX-Nemotron-3.5-Lightning-30B-A3B-MLX-AXQ-MXFP4-MTP/resolve/main/runtime_audit.json
2.89 kB
| { | |
| "applied_config_corrections": [], | |
| "current_config_sha256": "6a6d3246c4db834f09eae87deff809e157ca26617a23bded100de42a0ff5a1fa", | |
| "date": "2026-10-06", | |
| "header_sha256": { | |
| "model-00001-of-00004.safetensors": "ac490936b49fb441500f2f91d68dfe3c9d098559dad6303d024ba4e76c2c2b76", | |
| "model-00002-of-00004.safetensors": "b313f2840234aff995f100eef741083b5ae1af9d253661b76b8aabc4e42ca3b2", | |
| "model-00003-of-00004.safetensors": "eee80c2f961f201a8536e7a680944e9faaeb05d32c398f8778df6fd3e6f51650", | |
| "model-00004-of-00004.safetensors": "6028d48a118e3dd7bee2750d68937ea805510b73983af01ae61322e0c94f6b69", | |
| "mtp.safetensors": "dfc474d9015b22f6524ff0cb2a88fbad72dde8493cc169e4e3267ef444a78696" | |
| }, | |
| "input_bindings": { | |
| "ax_nemotron_mtp_manifest.json": "9338ce8221bd67be7172acf465172fdf0fd023722bd0034e579e9e97dd359af5", | |
| "axquant_manifest.json": "84a4060ece33dd6677e600ce32e942908c8ff98f1063303a433f8aa99e21a06d", | |
| "axquant_mtp_sidecar_manifest.json": "70868bd7b9910518b65bb4de2cd9d2119cb54db0d3b4db4c44a21b1b8233bf08", | |
| "config.json": "6a6d3246c4db834f09eae87deff809e157ca26617a23bded100de42a0ff5a1fa", | |
| "model.safetensors.index.json": "3e068938b04a3279ca3e6e1197882741cd5f67d5f7bce238a099f9cbb4da429e" | |
| }, | |
| "issues": [], | |
| "model_type": "nemotron_h", | |
| "mtp_files": [ | |
| "mtp.safetensors" | |
| ], | |
| "ngram_action": "none; do not invent n-gram data", | |
| "ngram_files": [], | |
| "ngram_quantization": [], | |
| "ngram_table_metadata": null, | |
| "ngram_tensor_count": 0, | |
| "quality_certified": false, | |
| "quantization": { | |
| "quantization": { | |
| "container_mode": "affine", | |
| "per_module_modes": { | |
| "affine": 1, | |
| "mxfp4": 162 | |
| }, | |
| "physical_recipe_issues": [] | |
| }, | |
| "quantization_config": { | |
| "container_mode": "affine", | |
| "per_module_modes": { | |
| "affine": 1, | |
| "mxfp4": 162 | |
| }, | |
| "physical_recipe_issues": [] | |
| } | |
| }, | |
| "removed_source_evidence": [], | |
| "repo_id": "AutomatosX/AX-Nemotron-3.5-Lightning-30B-A3B-MLX-AXQ-MXFP4-MTP", | |
| "runtime_arch_id": null, | |
| "runtime_verified": false, | |
| "schema_version": "axquant.hub-runtime-audit.v1", | |
| "scope": "Pinned remote config/index/Safetensors header audit; no runtime load or generation claim.", | |
| "source_revision": "5ab1716c9885e9bce30c063d159f5da4e0bb8c6f", | |
| "status": "no-audited-peer-export-profile", | |
| "unchanged_weight_sha256": { | |
| "model-00001-of-00004.safetensors": "a8b6891cdb0d7a97316fe841c48fc0b926588a555ecd5b08ed471e7468d58f51", | |
| "model-00002-of-00004.safetensors": "033922fe5c61923583d6cd8e72028b5d5fb1c5b0e788556356492581c7b7d70f", | |
| "model-00003-of-00004.safetensors": "35999dfcfdc58c5084e50df422b1b5ad6fb5a2f80c643f159c61fd4cd5ab75d8", | |
| "model-00004-of-00004.safetensors": "3c28f34687c69b0011ac52de0c557540d967d8df37442ebddc69dc71027c5c03", | |
| "mtp.safetensors": "40dd606b285acd5045bd4246a916463eeb1d5f4cc7a4ebe051555985ec436e72" | |
| } | |
| } | |