File size: 606 Bytes
cf231a1
 
 
 
 
 
 
 
 
 
 
 
a712daf
cf231a1
 
 
 
 
a712daf
 
 
cf231a1
a712daf
 
 
 
cf231a1
 
 
a712daf
1
2
3
4
5
6
7
8
9
10
11
12
13
14
15
16
17
18
19
20
21
22
23
24
25
26
27
28
29
30
31
---
license: apache-2.0
base_model: ByteDance/Ouro-1.4B-Thinking
library_name: transformers
pipeline_tag: text-generation
tags:
  - looped-language-model
  - recurrent-depth
  - sft
  - math
---

# Usage

```python
from transformers import AutoModelForCausalLM, AutoTokenizer

REPO = "omar81939/Ouro-1.4B-Thinking-depth-SFT"
DEPTH = 16

tokenizer = AutoTokenizer.from_pretrained(REPO)
model = AutoModelForCausalLM.from_pretrained(
    REPO,
    trust_remote_code=True,
    dtype="bfloat16",
    total_ut_steps=DEPTH,
)
```

With vLLM, set `hf_overrides={"total_ut_steps": DEPTH}` when creating the engine.