Instructions to use allenai/Olmo-Hybrid-Think-SFT-7B with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use allenai/Olmo-Hybrid-Think-SFT-7B with Transformers:
# Use a pipeline as a high-level helper from transformers import pipeline pipe = pipeline("text-generation", model="allenai/Olmo-Hybrid-Think-SFT-7B") messages = [ {"role": "user", "content": "Who are you?"}, ] pipe(messages)# pip install -U transformers accelerate # Load model directly from transformers import AutoTokenizer, AutoModelForCausalLM tokenizer = AutoTokenizer.from_pretrained("allenai/Olmo-Hybrid-Think-SFT-7B") model = AutoModelForCausalLM.from_pretrained("allenai/Olmo-Hybrid-Think-SFT-7B", device_map="auto") messages = [ {"role": "user", "content": "Who are you?"}, ] inputs = tokenizer.apply_chat_template( messages, add_generation_prompt=True, tokenize=True, return_dict=True, return_tensors="pt", ).to(model.device) outputs = model.generate(**inputs, max_new_tokens=256) print(tokenizer.decode(outputs[0][inputs["input_ids"].shape[-1]:])) - Notebooks
- Google Colab
- Kaggle
- Local Apps Settings
- vLLM
How to use allenai/Olmo-Hybrid-Think-SFT-7B with vLLM:
Install from pip and serve model
# Install vLLM from pip: pip install vllm # Start the vLLM server: vllm serve "allenai/Olmo-Hybrid-Think-SFT-7B" # Call the server using curl (OpenAI-compatible API): curl -X POST "http://localhost:8000/v1/chat/completions" \ -H "Content-Type: application/json" \ --data '{ "model": "allenai/Olmo-Hybrid-Think-SFT-7B", "messages": [ { "role": "user", "content": "What is the capital of France?" } ] }'Use Docker
docker model run hf.co/allenai/Olmo-Hybrid-Think-SFT-7B
- SGLang
How to use allenai/Olmo-Hybrid-Think-SFT-7B with SGLang:
Install from pip and serve model
# Install SGLang from pip: pip install sglang # Start the SGLang server: python3 -m sglang.launch_server \ --model-path "allenai/Olmo-Hybrid-Think-SFT-7B" \ --host 0.0.0.0 \ --port 30000 # Call the server using curl (OpenAI-compatible API): curl -X POST "http://localhost:30000/v1/chat/completions" \ -H "Content-Type: application/json" \ --data '{ "model": "allenai/Olmo-Hybrid-Think-SFT-7B", "messages": [ { "role": "user", "content": "What is the capital of France?" } ] }'Use Docker images
docker run --gpus all \ --shm-size 32g \ -p 30000:30000 \ -v ~/.cache/huggingface:/root/.cache/huggingface \ --env "HF_TOKEN=<secret>" \ --ipc=host \ lmsysorg/sglang:latest \ python3 -m sglang.launch_server \ --model-path "allenai/Olmo-Hybrid-Think-SFT-7B" \ --host 0.0.0.0 \ --port 30000 # Call the server using curl (OpenAI-compatible API): curl -X POST "http://localhost:30000/v1/chat/completions" \ -H "Content-Type: application/json" \ --data '{ "model": "allenai/Olmo-Hybrid-Think-SFT-7B", "messages": [ { "role": "user", "content": "What is the capital of France?" } ] }' - Docker Model Runner
How to use allenai/Olmo-Hybrid-Think-SFT-7B with Docker Model Runner:
docker model run hf.co/allenai/Olmo-Hybrid-Think-SFT-7B
Update README.md
Browse files
README.md
CHANGED
|
@@ -130,22 +130,22 @@ Moo Moo the cow would certinaly win.
|
|
| 130 |
|
| 131 |
| Skill | Benchmark | **Olmo Hybrid Think SFT 7B** | Olmo 3 Think 7B SFT | Olmo 3 Think 7B DPO | Olmo 3 Think 7B | OpenThinker3-7B | Nemotron-Nano-9B-v2 | DeepSeek-R1-Distill-Qwen-7B | Qwen 3 8B (reasoning) | Qwen 3 VL 8B Thinker | OpenReasoning Nemotron 7B |
|
| 132 |
|-------|-----------|--------------------------|------------------|------------------|--------------|------------------|-----------------------|------------------------------|-------------------------|---------------------------|-----------------------------|
|
| 133 |
-
| **Math** | MATH |
|
| 134 |
-
| | AIME 2024 |
|
| 135 |
-
| | AIME 2025 |
|
| 136 |
-
| | OMEGA |
|
| 137 |
-
| **Reasoning** | BBH |
|
| 138 |
-
| | ZebraLogic |
|
| 139 |
| | AGI Eval | | 77.2 | 79.1 | 81.5 | 78.6 | 83.1 | 69.5 | 87.0 | 90.1 | 81.4 |
|
| 140 |
-
| **Coding** | HumanEval+ |
|
| 141 |
-
| | MBPP+ |
|
| 142 |
-
| | LCB v3 |
|
| 143 |
-
| **IF** | IFEval |
|
| 144 |
-
| | IFBench |
|
| 145 |
-
| **Knowledge** | MMLU |
|
| 146 |
-
| **QA** | PopQA |
|
| 147 |
-
| | GPQA |
|
| 148 |
-
| **Chat** | AE 2 |
|
| 149 |
| **Safety** | | | 65.8 | 67.7 | 70.7 | 31.3 | 72.1 | 54.0 | 68.3 | 82.9 | 30.3 |
|
| 150 |
|
| 151 |
## Model Details
|
|
|
|
| 130 |
|
| 131 |
| Skill | Benchmark | **Olmo Hybrid Think SFT 7B** | Olmo 3 Think 7B SFT | Olmo 3 Think 7B DPO | Olmo 3 Think 7B | OpenThinker3-7B | Nemotron-Nano-9B-v2 | DeepSeek-R1-Distill-Qwen-7B | Qwen 3 8B (reasoning) | Qwen 3 VL 8B Thinker | OpenReasoning Nemotron 7B |
|
| 132 |
|-------|-----------|--------------------------|------------------|------------------|--------------|------------------|-----------------------|------------------------------|-------------------------|---------------------------|-----------------------------|
|
| 133 |
+
| **Math** | MATH | 93.8 | 94.4 | 92.4 | 95.1 | 94.5 | 94.4 | 87.9 | 95.1 | 95.2 | 94.6 |
|
| 134 |
+
| | AIME 2024 | 66.2 | 69.6 | 74.6 | 71.6 | 67.7 | 72.1 | 54.9 | 74.0 | 70.9 | 77.0 |
|
| 135 |
+
| | AIME 2025 | 55.2 | 57.6 | 62.7 | 64.6 | 57.2 | 58.9 | 40.2 | 67.8 | 61.5 | 73.1 |
|
| 136 |
+
| | OMEGA | 35.1 | 45.0 | 40.5 | 37.8 | 38.4 | 42.4 | 28.5 | 43.4 | 38.1 | 43.2 |
|
| 137 |
+
| **Reasoning** | BBH | 84.6 | 84.1 | 83.7 | 86.6 | 77.1 | 86.2 | 73.5 | 84.4 | 86.8 | 81.3 |
|
| 138 |
+
| | ZebraLogic | 55.1 | 57.9 | 60.6 | 66.5 | 34.9 | 60.8 | 26.1 | 85.2 | 91.2 | 22.4 |
|
| 139 |
| | AGI Eval | | 77.2 | 79.1 | 81.5 | 78.6 | 83.1 | 69.5 | 87.0 | 90.1 | 81.4 |
|
| 140 |
+
| **Coding** | HumanEval+ | 86.3 | 88.2 | 91.4 | 89.9 | 87.4 | 89.7 | 83.0 | 80.2 | 83.7 | 89.7 |
|
| 141 |
+
| | MBPP+ | 63.7 | 63.2 | 63.0 | 64.7 | 61.4 | 66.1 | 63.5 | 69.1 | 63.0 | 61.2 |
|
| 142 |
+
| | LCB v3 | 65.5 | 67.8 | 75.1 | 75.2 | 68.0 | 83.4 | 58.8 | 86.2 | 85.5 | 82.3 |
|
| 143 |
+
| **IF** | IFEval | 80.4 | 77.9 | 75.9 | 88.2 | 51.7 | 86.0 | 59.6 | 87.4 | 85.5 | 42.5 |
|
| 144 |
+
| | IFBench | 31.6 | 30.0 | 28.3 | 41.6 | 23.0 | 34.6 | 16.7 | 37.1 | 40.4 | 23.4 |
|
| 145 |
+
| **Knowledge** | MMLU | 80.5 | 74.9 | 74.8 | 77.8 | 77.4 | 84.3 | 67.9 | 85.4 | 86.5 | 80.7 |
|
| 146 |
+
| **QA** | PopQA | 25.1 | 20.8 | 24.7 | 23.7 | 18.0 | 17.9 | 12.8 | 24.3 | 29.3 | 14.5 |
|
| 147 |
+
| | GPQA | 47.0 | 45.8 | 48.6 | 46.2 | 47.6 | 56.2 | 54.4 | 57.7 | 61.5 | 56.6 |
|
| 148 |
+
| **Chat** | AE 2 | 49.0 | 43.9 | 50.6 | 52.1 | 24.0 | 58.0 | 7.7 | 60.5 | 73.5 | 8.6 |
|
| 149 |
| **Safety** | | | 65.8 | 67.7 | 70.7 | 31.3 | 72.1 | 54.0 | 68.3 | 82.9 | 30.3 |
|
| 150 |
|
| 151 |
## Model Details
|