SAnocha commited on
Commit
c6f69e5
·
verified ·
1 Parent(s): e803b65

Update README.md

Browse files
Files changed (1) hide show
  1. README.md +1 -0
README.md CHANGED
@@ -119,6 +119,7 @@ Note:
119
 
120
  The training counts and dataset mixture details reported in this model card reflect the exact constructed training pool consumed during this specific model run. Figures may differ slightly from public dataset releases, which represent downloadable open-source subsets of the broader corpus.
121
 
 
122
 
123
  ## Call for Contributions
124
  We encourage researchers, developers, and language enthusiasts to actively contribute to the enhancement and expansion of SEA-LION. Contributions can involve identifying and reporting bugs, sharing pre-training, instruction, and preference data, improving documentation usability, proposing and implementing new model evaluation tasks and metrics, or training versions of the model in additional Southeast Asian languages. Join us in shaping the future of SEA-LION by sharing your expertise and insights to make these models more accessible, accurate, and versatile. Please check out our GitHub for further information on the call for contributions.
 
119
 
120
  The training counts and dataset mixture details reported in this model card reflect the exact constructed training pool consumed during this specific model run. Figures may differ slightly from public dataset releases, which represent downloadable open-source subsets of the broader corpus.
121
 
122
+ This model card serves as an immutable record of the final released checkpoint. Consequently, the hardware specifications, compute hours, and precise dataset volumes (such as final filtered instruction counts) reported here reflect the exact production run used to generate this specific artifact. These figures may differ from the aggregate totals, pre-filtered data pools, or preliminary experimental runs (e.g., initial H100 benchmarks) documented in our accompanying research papers.
123
 
124
  ## Call for Contributions
125
  We encourage researchers, developers, and language enthusiasts to actively contribute to the enhancement and expansion of SEA-LION. Contributions can involve identifying and reporting bugs, sharing pre-training, instruction, and preference data, improving documentation usability, proposing and implementing new model evaluation tasks and metrics, or training versions of the model in additional Southeast Asian languages. Join us in shaping the future of SEA-LION by sharing your expertise and insights to make these models more accessible, accurate, and versatile. Please check out our GitHub for further information on the call for contributions.