chen11003 commited on
Commit
82f8dd2
·
verified ·
1 Parent(s): 87782f8

Update README.md

Browse files
Files changed (1) hide show
  1. README.md +0 -15
README.md CHANGED
@@ -212,21 +212,6 @@ submetric over the pre-distillation merge. This makes the model useful for
212
  research on compact reasoning models, local inference, distillation, and
213
  task-specific adaptation.
214
 
215
- ## Model Card Comparison Table
216
-
217
- | Benchmark | **K2-Horizon-0.9B** | MiniCPM5-1B | Qwen3.5-0.8B | Qwen3.5-2B |
218
- |---|---:|---:|---:|---:|
219
- | IFEval (strict instruction) | **80.8†** | 80.41‡ | 44.0‡ | 78.6‡ |
220
- | GPQA-Diamond (avg@16) | **27.3†** | 26.26‡ | 11.9‡ | 51.6‡ |
221
- | HMMT February 2026 (avg@16) | **25.8†** | 23.3† | 0.57‡ | 18.56† |
222
- | AIME 2025 (avg@16) | **41.7†** | 40.42‡ | 1.04‡ | 26.46† |
223
- | AIME 2026 (avg@16) | **48.5†** | 40.42‡ | 0.21‡ | 25.42† |
224
- | HumanEval+ (pass@1) | **79.9†** | 65.2† | 26.22† | 42.68† |
225
- | MBPP+ (pass@1) | **68.0†** | 60.6† | 32.8† | 47.09† |
226
- | LiveCodeBench v6 (avg@3) | **37.41†** | 33.52‡ | 5.33‡ | 13.08† |
227
-
228
- - **† Local result.**
229
- - **‡ Published comparison/model-card value; protocol is not necessarily matched.**
230
 
231
  ## How to Use
232
 
 
212
  research on compact reasoning models, local inference, distillation, and
213
  task-specific adaptation.
214
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
215
 
216
  ## How to Use
217