Original model description:

license: apache-2.0 library_name: transformers tags: - mergekit - merge base_model: - elinas/Chronos-Gold-12B-1.0 - Triangle104/ChatWaifu_Magnum_V0.2 model-index: - name: Chatty-Harry_V3.0 results: - task: type: text-generation name: Text Generation dataset: name: IFEval (0-Shot) type: HuggingFaceH4/ifeval args: num_few_shot: 0 metrics: - type: inst_level_strict_acc and prompt_level_strict_acc value: 36.75 name: strict accuracy source: url: https://huggingface.co/spaces/open-llm-leaderboard/open_llm_leaderboard?query=Triangle104/Chatty-Harry_V3.0 name: Open LLM Leaderboard - task: type: text-generation name: Text Generation dataset: name: BBH (3-Shot) type: BBH args: num_few_shot: 3 metrics: - type: acc_norm value: 35.89 name: normalized accuracy source: url: https://huggingface.co/spaces/open-llm-leaderboard/open_llm_leaderboard?query=Triangle104/Chatty-Harry_V3.0 name: Open LLM Leaderboard - task: type: text-generation name: Text Generation dataset: name: MATH Lvl 5 (4-Shot) type: hendrycks/competition_math args: num_few_shot: 4 metrics: - type: exact_match value: 10.57 name: exact match source: url: https://huggingface.co/spaces/open-llm-leaderboard/open_llm_leaderboard?query=Triangle104/Chatty-Harry_V3.0 name: Open LLM Leaderboard - task: type: text-generation name: Text Generation dataset: name: GPQA (0-shot) type: Idavidrein/gpqa args: num_few_shot: 0 metrics: - type: acc_norm value: 9.73 name: acc_norm source: url: https://huggingface.co/spaces/open-llm-leaderboard/open_llm_leaderboard?query=Triangle104/Chatty-Harry_V3.0 name: Open LLM Leaderboard - task: type: text-generation name: Text Generation dataset: name: MuSR (0-shot) type: TAUR-Lab/MuSR args: num_few_shot: 0 metrics: - type: acc_norm value: 15.04 name: acc_norm source: url: https://huggingface.co/spaces/open-llm-leaderboard/open_llm_leaderboard?query=Triangle104/Chatty-Harry_V3.0 name: Open LLM Leaderboard - task: type: text-generation name: Text Generation dataset: name: MMLU-PRO (5-shot) type: TIGER-Lab/MMLU-Pro config: main split: test args: num_few_shot: 5 metrics: - type: acc value: 30.02 name: accuracy source: url: https://huggingface.co/spaces/open-llm-leaderboard/open_llm_leaderboard?query=Triangle104/Chatty-Harry_V3.0 name: Open LLM Leaderboard

Model details:

You've got to ask yourself one question...

Added some Chronos-Gold-12B-1.0 at suggestion.

Feedback is welcome.

merge

This is a merge of pre-trained language models created using mergekit.

Merge Details

Merge Method

This model was merged using the TIES merge method using Triangle104/ChatWaifu_Magnum_V0.2 as a base.

Models Merged

The following models were included in the merge:

elinas/Chronos-Gold-12B-1.0

Configuration

The following YAML configuration was used to produce this model:

models:
  - model: elinas/Chronos-Gold-12B-1.0
    parameters:
      density: 0.5
      weight: 0.5

merge_method: ties
base_model: Triangle104/ChatWaifu_Magnum_V0.2
parameters:
  normalize: false
  int8_mask: true
dtype: float16

Open LLM Leaderboard Evaluation Results

Detailed results can be found here

Metric	Value
Avg.	23.00
IFEval (0-Shot)	36.75
BBH (3-Shot)	35.89
MATH Lvl 5 (4-Shot)	10.57
GPQA (0-shot)	9.73
MuSR (0-shot)	15.04
MMLU-PRO (5-shot)	30.02

Downloads last month: 2

Safetensors

Model size

12B params

Tensor type

F32

F16

Inference Providers NEW

This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Paper for RichardErkhov/Triangle104_-_Chatty-Harry_V3.0-4bits

Resolving Interference When Merging Models

Paper • 2306.01708 • Published Jun 2, 2023 • 19