Maggio33 commited on
Commit
abd5b04
·
verified ·
1 Parent(s): 117ebb3

Card: 64M 560k-600k, 128M 380k (progress rows only)

Browse files
Files changed (1) hide show
  1. README.md +4 -0
README.md CHANGED
@@ -191,6 +191,9 @@ Two final models are training now with a recipe fixed before the runs started. N
191
  | 64M final (14×576) | 500,000 (65.8%) | 47.05 | 75.55 | 2.3693 | 75.41 | training in progress; not a result |
192
  | 64M final (14×576) | 520,000 (68.4%) | 46.80 | 75.60 | 2.3674 | 75.35 | training in progress; not a result |
193
  | 64M final (14×576) | 540,000 (71.1%) | 46.46 | 75.83 | 2.3678 | 75.31 | training in progress; not a result |
 
 
 
194
  | 128M final (16×768) | 20,000 (2.6%) | 43.90 | 73.91 | 2.5383 | 71.34 | training in progress; not a result |
195
  | 128M final (16×768) | 40,000 (5.3%) | 46.68 | 74.81 | 2.4612 | 72.77 | training in progress; not a result |
196
  | 128M final (16×768) | 60,000 (7.9%) | 47.43 | 75.24 | 2.4151 | 73.28 | training in progress; not a result |
@@ -209,6 +212,7 @@ Two final models are training now with a recipe fixed before the runs started. N
209
  | 128M final (16×768) | 320,000 (42.1%) | 50.46 | 77.79 | 2.3026 | 75.44 | training in progress; not a result |
210
  | 128M final (16×768) | 340,000 (44.7%) | 50.13 | 78.32 | 2.3040 | 75.50 | training in progress; not a result |
211
  | 128M final (16×768) | 360,000 (47.4%) | 50.84 | 78.01 | 2.3011 | 75.64 | training in progress; not a result |
 
212
 
213
  Reference, same protocol: the v1 models at the end of their runs (step 400,000) and at the same steps as the final runs.
214
 
 
191
  | 64M final (14×576) | 500,000 (65.8%) | 47.05 | 75.55 | 2.3693 | 75.41 | training in progress; not a result |
192
  | 64M final (14×576) | 520,000 (68.4%) | 46.80 | 75.60 | 2.3674 | 75.35 | training in progress; not a result |
193
  | 64M final (14×576) | 540,000 (71.1%) | 46.46 | 75.83 | 2.3678 | 75.31 | training in progress; not a result |
194
+ | 64M final (14×576) | 560,000 (73.7%) | 47.81 | 76.17 | 2.3648 | 75.90 | training in progress; not a result |
195
+ | 64M final (14×576) | 580,000 (76.3%) | 47.47 | 76.38 | 2.3604 | 75.87 | training in progress; not a result |
196
+ | 64M final (14×576) | 600,000 (78.9%) | 47.56 | 76.11 | 2.3561 | 75.81 | training in progress; not a result |
197
  | 128M final (16×768) | 20,000 (2.6%) | 43.90 | 73.91 | 2.5383 | 71.34 | training in progress; not a result |
198
  | 128M final (16×768) | 40,000 (5.3%) | 46.68 | 74.81 | 2.4612 | 72.77 | training in progress; not a result |
199
  | 128M final (16×768) | 60,000 (7.9%) | 47.43 | 75.24 | 2.4151 | 73.28 | training in progress; not a result |
 
212
  | 128M final (16×768) | 320,000 (42.1%) | 50.46 | 77.79 | 2.3026 | 75.44 | training in progress; not a result |
213
  | 128M final (16×768) | 340,000 (44.7%) | 50.13 | 78.32 | 2.3040 | 75.50 | training in progress; not a result |
214
  | 128M final (16×768) | 360,000 (47.4%) | 50.84 | 78.01 | 2.3011 | 75.64 | training in progress; not a result |
215
+ | 128M final (16×768) | 380,000 (50.0%) | 51.18 | 78.86 | 2.2994 | 76.05 | training in progress; not a result |
216
 
217
  Reference, same protocol: the v1 models at the end of their runs (step 400,000) and at the same steps as the final runs.
218