YOYO-AI commited on
Commit
7d16be9
·
verified ·
1 Parent(s): 3458f67

Update README.md

Browse files
Files changed (1) hide show
  1. README.md +111 -0
README.md CHANGED
@@ -12,6 +12,101 @@ base_model:
12
  pipeline_tag: text-generation
13
  tags:
14
  - merge
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
15
  ---
16
  ![image/png](https://cdn-uploads.huggingface.co/production/uploads/64e174e202fa032de4143324/8YkBIMWfWNXm0dbNwj2HH.png)
17
  # ZYH-LLM-Qwen2.5-14B-V3
@@ -19,6 +114,22 @@ This is the third-generation model of the **ZYH-LLM series**.
19
 
20
  It employs a large amount of model merging techniques, aiming to provide a **powerful and unified 14-billion-parameter model**, laying a solid foundation for further model merging and model fine-tuning.
21
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
22
  The following are the specific details of model merging, hoping to inspire you:
23
 
24
  ## First stage:
 
12
  pipeline_tag: text-generation
13
  tags:
14
  - merge
15
+ model-index:
16
+ - name: ZYH-LLM-Qwen2.5-14B-V3
17
+ results:
18
+ - task:
19
+ type: text-generation
20
+ name: Text Generation
21
+ dataset:
22
+ name: IFEval (0-Shot)
23
+ type: HuggingFaceH4/ifeval
24
+ args:
25
+ num_few_shot: 0
26
+ metrics:
27
+ - type: inst_level_strict_acc and prompt_level_strict_acc
28
+ value: 85.78
29
+ name: strict accuracy
30
+ source:
31
+ url: https://huggingface.co/spaces/open-llm-leaderboard/open_llm_leaderboard?query=YOYO-AI/ZYH-LLM-Qwen2.5-14B-V3
32
+ name: Open LLM Leaderboard
33
+ - task:
34
+ type: text-generation
35
+ name: Text Generation
36
+ dataset:
37
+ name: BBH (3-Shot)
38
+ type: BBH
39
+ args:
40
+ num_few_shot: 3
41
+ metrics:
42
+ - type: acc_norm
43
+ value: 48.18
44
+ name: normalized accuracy
45
+ source:
46
+ url: https://huggingface.co/spaces/open-llm-leaderboard/open_llm_leaderboard?query=YOYO-AI/ZYH-LLM-Qwen2.5-14B-V3
47
+ name: Open LLM Leaderboard
48
+ - task:
49
+ type: text-generation
50
+ name: Text Generation
51
+ dataset:
52
+ name: MATH Lvl 5 (4-Shot)
53
+ type: hendrycks/competition_math
54
+ args:
55
+ num_few_shot: 4
56
+ metrics:
57
+ - type: exact_match
58
+ value: 52.72
59
+ name: exact match
60
+ source:
61
+ url: https://huggingface.co/spaces/open-llm-leaderboard/open_llm_leaderboard?query=YOYO-AI/ZYH-LLM-Qwen2.5-14B-V3
62
+ name: Open LLM Leaderboard
63
+ - task:
64
+ type: text-generation
65
+ name: Text Generation
66
+ dataset:
67
+ name: GPQA (0-shot)
68
+ type: Idavidrein/gpqa
69
+ args:
70
+ num_few_shot: 0
71
+ metrics:
72
+ - type: acc_norm
73
+ value: 10.96
74
+ name: acc_norm
75
+ source:
76
+ url: https://huggingface.co/spaces/open-llm-leaderboard/open_llm_leaderboard?query=YOYO-AI/ZYH-LLM-Qwen2.5-14B-V3
77
+ name: Open LLM Leaderboard
78
+ - task:
79
+ type: text-generation
80
+ name: Text Generation
81
+ dataset:
82
+ name: MuSR (0-shot)
83
+ type: TAUR-Lab/MuSR
84
+ args:
85
+ num_few_shot: 0
86
+ metrics:
87
+ - type: acc_norm
88
+ value: 9.00
89
+ name: acc_norm
90
+ source:
91
+ url: https://huggingface.co/spaces/open-llm-leaderboard/open_llm_leaderboard?query=YOYO-AI/ZYH-LLM-Qwen2.5-14B-V3
92
+ name: Open LLM Leaderboard
93
+ - task:
94
+ type: text-generation
95
+ name: Text Generation
96
+ dataset:
97
+ name: MMLU-PRO (5-shot)
98
+ type: TIGER-Lab/MMLU-Pro
99
+ config: main
100
+ split: test
101
+ args:
102
+ num_few_shot: 5
103
+ metrics:
104
+ - type: acc
105
+ value: 43.12
106
+ name: accuracy
107
+ source:
108
+ url: https://huggingface.co/spaces/open-llm-leaderboard/open_llm_leaderboard?query=YOYO-AI/ZYH-LLM-Qwen2.5-14B-V3
109
+ name: Open LLM Leaderboard
110
  ---
111
  ![image/png](https://cdn-uploads.huggingface.co/production/uploads/64e174e202fa032de4143324/8YkBIMWfWNXm0dbNwj2HH.png)
112
  # ZYH-LLM-Qwen2.5-14B-V3
 
114
 
115
  It employs a large amount of model merging techniques, aiming to provide a **powerful and unified 14-billion-parameter model**, laying a solid foundation for further model merging and model fine-tuning.
116
 
117
+ # As of February 25, 2025, the 14B model with the highest IFEval score
118
+ ![image/png](https://cdn-uploads.huggingface.co/production/uploads/64e174e202fa032de4143324/e2dZMCmjF4utR07q9yjAx.png)
119
+
120
+ # [Open LLM Leaderboard Evaluation Results](https://huggingface.co/spaces/open-llm-leaderboard/open_llm_leaderboard)
121
+ Detailed results can be found [here](https://huggingface.co/datasets/open-llm-leaderboard/YOYO-AI__ZYH-LLM-Qwen2.5-14B-V3-details)
122
+
123
+ | Metric |Value|
124
+ |-------------------|----:|
125
+ |Avg. |41.63|
126
+ |IFEval (0-Shot) |85.78|
127
+ |BBH (3-Shot) |48.18|
128
+ |MATH Lvl 5 (4-Shot)|52.72|
129
+ |GPQA (0-shot) |10.96|
130
+ |MuSR (0-shot) |9.00|
131
+ |MMLU-PRO (5-shot) |43.12|
132
+
133
  The following are the specific details of model merging, hoping to inspire you:
134
 
135
  ## First stage: