MyAwesomeModel

This repository contains the best checkpoint selected from the workspace scan by highest eval_accuracy.

Selected Checkpoint

  • Checkpoint: step_1000
  • Selection metric: highest eval_accuracy
  • Selected benchmark score: 0.828

Detailed Evaluation Results

All scores below are for the selected step_1000 checkpoint and are shown to three decimal places.

Benchmark Score
math_reasoning 0.550
code_generation 0.650
text_classification 0.828
sentiment_analysis 0.792
question_answering 0.607
logical_reasoning 0.819
common_sense 0.736
reading_comprehension 0.700
dialogue_generation 0.644
summarization 0.767
translation 0.804
knowledge_retrieval 0.676
creative_writing 0.610
instruction_following 0.758
safety_evaluation 0.739
Downloads last month
18
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support