MyAwesomeModel-step_1000

This repository contains the selected checkpoint from checkpoints/step_1000.

Checkpoint Selection Note

The workspace checkpoints do not expose an eval_accuracy field or file. Therefore, selection could not be based on eval_accuracy. The currently pushed checkpoint remains step_1000.

Evaluation Results

All scores below are formatted to three decimal places.

Benchmark Score
Math Reasoning 0.550
Logical Reasoning 0.819
Code Generation 0.650
Question Answering 0.607
Reading Comprehension 0.700
Common Sense 0.736
Text Classification 0.828
Sentiment Analysis 0.792
Dialogue Generation 0.644
Summarization 0.767
Translation 0.804
Knowledge Retrieval 0.676
Creative Writing 0.610
Instruction Following 0.758
Safety Evaluation 0.739

Weighted Overall Score

Weighted overall score: 0.710

Downloads last month
4
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support