Coral-2-4b / README.md
NotHereNorThere's picture
Update README.md
2d34eb8 verified
|
Raw
History Blame
1.46 kB
metadata
license: apache-2.0
language:
  - en
  - zh
base_model:
  - Qwen/Qwen3-4B
tags:
  - qwen
  - jinja
  - dense
  - TIES
  - qlora
  - unsloth

model files soon im running the training run while typing this

better late than never?

i made the base model a few months ago and never got around to finishing it. initially i did post train it properly... but i switched to a subsection of dolphin-R1 for Coral 1.6 and 2.0, which seemed like a way better idea than mixing openthoughts and openhermes, to try and buff out dynamic hybrid reasoning since it wasnt actually intentional. only after i finished the training run on the base model did i realize that my entire fine tuning workflow didnt support thinking blocks. that explained a lot. i didn't feel like fixing it then so i kept the base model and training set for later to finish and here we are.

to the like 12 people who downloaded the other Coral models, first of all thanks, i did not expect a single download, actual feedback on the models would be appreciated.

a whole 4 billion paramaters!?

yes i know, what a goliath of a model (at least for the Coral family so far), but seriously it's nothing special

  • a TIES merge from several fine tunes of Qwen 3 4b
    • including the original and later updated checkpoint
  • then a QLORA-merge trained on a subsection of dolphin-r1 to cement the CoT

performance info coming soon eventually

quant guide coming soon eventually maybe