NoYo25 commited on
Commit
d40ab2d
·
1 Parent(s): 87f4de8

Add pre-training hyperparams

Browse files
Files changed (1) hide show
  1. README.md +7 -0
README.md CHANGED
@@ -36,6 +36,13 @@ training_data:
36
  - corpora:
37
  - (+Abs) Springer and Elsevier abstracts in the duration of 1990-2020
38
  - (+Abs+Full) Springer and Elsevier abstracts and open access full publication text in the duration of 1990-2020
 
 
 
 
 
 
 
39
  ---
40
 
41
  # BiodivBERT
 
36
  - corpora:
37
  - (+Abs) Springer and Elsevier abstracts in the duration of 1990-2020
38
  - (+Abs+Full) Springer and Elsevier abstracts and open access full publication text in the duration of 1990-2020
39
+ pre-training-hyperparams:
40
+ - MAX_LEN = 512 # Default of BERT Tokenizer
41
+ - MLM_PROP = 0.15 # Data Collator
42
+ - num_train_epochs = 3 # the minimum sufficient epochs found on many articles && default of trainer here
43
+ - per_device_train_batch_size = 16 # the maximumn that could be held by V100 on Ara with 512 MAX_LEN was 8 in the old run
44
+ - per_device_eval_batch_size = 16 # usually as above
45
+ - gradient_accumulation_steps = 4 # this will grant a minim batch size 16 * 4 * nGPUs.
46
  ---
47
 
48
  # BiodivBERT