SeaWolf-AI commited on
Commit
bc3c7b0
·
verified ·
1 Parent(s): 0e330b7

card: clarify Model-level RSI wording

Browse files
Files changed (1) hide show
  1. README.md +1 -1
README.md CHANGED
@@ -226,7 +226,7 @@ Practice sets are deduplicated against every evaluation set we report (8-gram ov
226
 
227
  ### Model-level RSI vs. harness-level RSI
228
 
229
- Darwin-180B-RSI is **Model-level RSI**: the model itself (its weights) improves, with no new human-written answers. The answer keys used for verification come from existing public datasets. **Harness-level RSI** (e.g., Google's RRSI) improves the prompts, tools and workflow around a fixed model. It's like rewriting an employee's manual, while Model-level RSI is the employee getting smarter. The two are complementary.
230
 
231
  ---
232
 
 
226
 
227
  ### Model-level RSI vs. harness-level RSI
228
 
229
+ Darwin-180B-RSI is **Model-level RSI**: the model itself (its weights) improves by learning only from its own solutions. No human-written solutions or reasoning traces are used; correctness is checked automatically. **Harness-level RSI** (e.g., Google's RRSI) improves the prompts, tools and workflow around a fixed model. It's like rewriting an employee's manual, while Model-level RSI is the employee getting smarter. The two are complementary.
230
 
231
  ---
232