aashish1904 commited on
Commit
4183509
1 Parent(s): 4c264ad

Upload README.md with huggingface_hub

Browse files
Files changed (1) hide show
  1. README.md +171 -0
README.md ADDED
@@ -0,0 +1,171 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+
2
+ ---
3
+
4
+ license: llama3
5
+ library_name: transformers
6
+ tags:
7
+ - mergekit
8
+ - merge
9
+ base_model:
10
+ - flammenai/Mahou-1.2-llama3-8B
11
+ - lemon07r/Llama-3-RedMagic2-8B
12
+ - lemon07r/Lllama-3-RedElixir-8B
13
+ - NousResearch/Meta-Llama-3-8B
14
+ - nbeerbower/llama-3-spicy-abliterated-stella-8B
15
+ model-index:
16
+ - name: Llama-3-RedMagic4-8B
17
+ results:
18
+ - task:
19
+ type: text-generation
20
+ name: Text Generation
21
+ dataset:
22
+ name: IFEval (0-Shot)
23
+ type: HuggingFaceH4/ifeval
24
+ args:
25
+ num_few_shot: 0
26
+ metrics:
27
+ - type: inst_level_strict_acc and prompt_level_strict_acc
28
+ value: 48.64
29
+ name: strict accuracy
30
+ source:
31
+ url: https://huggingface.co/spaces/open-llm-leaderboard/open_llm_leaderboard?query=lemon07r/Llama-3-RedMagic4-8B
32
+ name: Open LLM Leaderboard
33
+ - task:
34
+ type: text-generation
35
+ name: Text Generation
36
+ dataset:
37
+ name: BBH (3-Shot)
38
+ type: BBH
39
+ args:
40
+ num_few_shot: 3
41
+ metrics:
42
+ - type: acc_norm
43
+ value: 19.48
44
+ name: normalized accuracy
45
+ source:
46
+ url: https://huggingface.co/spaces/open-llm-leaderboard/open_llm_leaderboard?query=lemon07r/Llama-3-RedMagic4-8B
47
+ name: Open LLM Leaderboard
48
+ - task:
49
+ type: text-generation
50
+ name: Text Generation
51
+ dataset:
52
+ name: MATH Lvl 5 (4-Shot)
53
+ type: hendrycks/competition_math
54
+ args:
55
+ num_few_shot: 4
56
+ metrics:
57
+ - type: exact_match
58
+ value: 8.31
59
+ name: exact match
60
+ source:
61
+ url: https://huggingface.co/spaces/open-llm-leaderboard/open_llm_leaderboard?query=lemon07r/Llama-3-RedMagic4-8B
62
+ name: Open LLM Leaderboard
63
+ - task:
64
+ type: text-generation
65
+ name: Text Generation
66
+ dataset:
67
+ name: GPQA (0-shot)
68
+ type: Idavidrein/gpqa
69
+ args:
70
+ num_few_shot: 0
71
+ metrics:
72
+ - type: acc_norm
73
+ value: 5.37
74
+ name: acc_norm
75
+ source:
76
+ url: https://huggingface.co/spaces/open-llm-leaderboard/open_llm_leaderboard?query=lemon07r/Llama-3-RedMagic4-8B
77
+ name: Open LLM Leaderboard
78
+ - task:
79
+ type: text-generation
80
+ name: Text Generation
81
+ dataset:
82
+ name: MuSR (0-shot)
83
+ type: TAUR-Lab/MuSR
84
+ args:
85
+ num_few_shot: 0
86
+ metrics:
87
+ - type: acc_norm
88
+ value: 4.38
89
+ name: acc_norm
90
+ source:
91
+ url: https://huggingface.co/spaces/open-llm-leaderboard/open_llm_leaderboard?query=lemon07r/Llama-3-RedMagic4-8B
92
+ name: Open LLM Leaderboard
93
+ - task:
94
+ type: text-generation
95
+ name: Text Generation
96
+ dataset:
97
+ name: MMLU-PRO (5-shot)
98
+ type: TIGER-Lab/MMLU-Pro
99
+ config: main
100
+ split: test
101
+ args:
102
+ num_few_shot: 5
103
+ metrics:
104
+ - type: acc
105
+ value: 29.73
106
+ name: accuracy
107
+ source:
108
+ url: https://huggingface.co/spaces/open-llm-leaderboard/open_llm_leaderboard?query=lemon07r/Llama-3-RedMagic4-8B
109
+ name: Open LLM Leaderboard
110
+
111
+ ---
112
+
113
+ ![](https://lh7-rt.googleusercontent.com/docsz/AD_4nXeiuCm7c8lEwEJuRey9kiVZsRn2W-b4pWlu3-X534V3YmVuVc2ZL-NXg2RkzSOOS2JXGHutDuyyNAUtdJI65jGTo8jT9Y99tMi4H4MqL44Uc5QKG77B0d6-JfIkZHFaUA71-RtjyYZWVIhqsNZcx8-OMaA?key=xt3VSDoCbmTY7o-cwwOFwQ)
114
+
115
+ # QuantFactory/Llama-3-RedMagic4-8B-GGUF
116
+ This is quantized version of [lemon07r/Llama-3-RedMagic4-8B](https://huggingface.co/lemon07r/Llama-3-RedMagic4-8B) created using llama.cpp
117
+
118
+ # Original Model Card
119
+
120
+ # Llama-3-RedMagic4-8B
121
+
122
+ This is a merge of pre-trained language models created using [mergekit](https://github.com/cg123/mergekit).
123
+
124
+ ## Merge Details
125
+ ### Merge Method
126
+
127
+ This model was merged using the [Model Stock](https://arxiv.org/abs/2403.19522) merge method using [NousResearch/Meta-Llama-3-8B](https://huggingface.co/NousResearch/Meta-Llama-3-8B) as a base.
128
+
129
+ ### Models Merged
130
+
131
+ The following models were included in the merge:
132
+ * [flammenai/Mahou-1.2-llama3-8B](https://huggingface.co/flammenai/Mahou-1.2-llama3-8B)
133
+ * [lemon07r/Llama-3-RedMagic2-8B](https://huggingface.co/lemon07r/Llama-3-RedMagic2-8B)
134
+ * [lemon07r/Lllama-3-RedElixir-8B](https://huggingface.co/lemon07r/Lllama-3-RedElixir-8B)
135
+ * [nbeerbower/llama-3-spicy-abliterated-stella-8B](https://huggingface.co/nbeerbower/llama-3-spicy-abliterated-stella-8B)
136
+
137
+ ### Configuration
138
+
139
+ The following YAML configuration was used to produce this model:
140
+
141
+ ```yaml
142
+ base_model: NousResearch/Meta-Llama-3-8B
143
+ dtype: bfloat16
144
+ merge_method: model_stock
145
+ slices:
146
+ - sources:
147
+ - layer_range: [0, 32]
148
+ model: lemon07r/Llama-3-RedMagic2-8B
149
+ - layer_range: [0, 32]
150
+ model: lemon07r/Lllama-3-RedElixir-8B
151
+ - layer_range: [0, 32]
152
+ model: nbeerbower/llama-3-spicy-abliterated-stella-8B
153
+ - layer_range: [0, 32]
154
+ model: flammenai/Mahou-1.2-llama3-8B
155
+ - layer_range: [0, 32]
156
+ model: NousResearch/Meta-Llama-3-8B
157
+ ```
158
+ # [Open LLM Leaderboard Evaluation Results](https://huggingface.co/spaces/open-llm-leaderboard/open_llm_leaderboard)
159
+ Detailed results can be found [here](https://huggingface.co/datasets/open-llm-leaderboard/details_lemon07r__Llama-3-RedMagic4-8B)
160
+
161
+ | Metric |Value|
162
+ |-------------------|----:|
163
+ |Avg. |19.32|
164
+ |IFEval (0-Shot) |48.64|
165
+ |BBH (3-Shot) |19.48|
166
+ |MATH Lvl 5 (4-Shot)| 8.31|
167
+ |GPQA (0-shot) | 5.37|
168
+ |MuSR (0-shot) | 4.38|
169
+ |MMLU-PRO (5-shot) |29.73|
170
+
171
+