Text Classification
Transformers
Safetensors
modernbert
agent-safety
tool-calling
long-context
distillation
Eval Results (legacy)
text-embeddings-inference
Instructions to use ProCreations/auto-200m-2 with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use ProCreations/auto-200m-2 with Transformers:
# Use a pipeline as a high-level helper from transformers import pipeline pipe = pipeline("text-classification", model="ProCreations/auto-200m-2")# Load model directly from transformers import AutoTokenizer, AutoModelForSequenceClassification tokenizer = AutoTokenizer.from_pretrained("ProCreations/auto-200m-2") model = AutoModelForSequenceClassification.from_pretrained("ProCreations/auto-200m-2", device_map="auto") - Notebooks
- Google Colab
- Kaggle
File size: 1,314 Bytes
0bbb929 | 1 2 3 4 5 6 7 8 9 10 11 12 13 14 15 16 17 18 19 20 21 22 23 24 25 26 27 28 29 30 31 32 33 34 35 36 37 38 39 40 41 42 43 44 45 46 47 48 49 50 51 52 53 54 | {
"results": [
{
"theta": 160000.0,
"short": 0.9160820834562606,
"mid": 1.0387379115148965,
"long": 1.1367697905727685,
"seconds": 26.022907972335815
},
{
"theta": 640000.0,
"short": 0.9236352910098977,
"mid": 1.0243582937320355,
"long": 0.9822827437326973,
"seconds": 11.011847972869873
},
{
"theta": 1280000.0,
"short": 0.9304989230334977,
"mid": 1.031460880067349,
"long": 0.9601743056856983,
"seconds": 11.087600231170654
},
{
"theta": 2560000.0,
"short": 0.939639166783192,
"mid": 1.044918367936274,
"long": 0.9650228895573634,
"seconds": 11.14451551437378
},
{
"theta": 5120000.0,
"short": 0.9498286157172525,
"mid": 1.0788166451562344,
"long": 1.0147006089168509,
"seconds": 11.193717956542969
},
{
"theta": 10240000.0,
"short": 0.9709671184858385,
"mid": 1.0868880858439343,
"long": 0.9965657126456587,
"seconds": 11.231298685073853
}
],
"chosen_theta": 1280000.0,
"samples": {
"short": 256,
"mid": 96,
"long": 64
},
"rule": "min(mid + long masked-LM loss) with short-context loss within 5% of the original 160k theta; 15% masking; training text only"
}
|