tugot17 iamleonie commited on
Commit
9235d67
Β·
1 Parent(s): 8a7cd88

Rename GGUF files: remove Draft-v1 from filenames (#1)

Browse files

- Rename GGUF files: remove Draft-v1 from filenames (a848f4d49defaf9c3953e968bdd6f9f52a126b3f)


Co-authored-by: Leonie Monigatti <iamleonie@users.noreply.huggingface.co>

.gitattributes CHANGED
@@ -36,3 +36,6 @@ saved_model/**/* filter=lfs diff=lfs merge=lfs -text
36
  LFM2.5-1.2B-Instruct-DSpark-Draft-v1-F16.gguf filter=lfs diff=lfs merge=lfs -text
37
  LFM2.5-1.2B-Instruct-DSpark-Draft-v1-Q8_0.gguf filter=lfs diff=lfs merge=lfs -text
38
  LFM2.5-1.2B-Instruct-DSpark-Draft-v1-Q4_K_M.gguf filter=lfs diff=lfs merge=lfs -text
 
 
 
 
36
  LFM2.5-1.2B-Instruct-DSpark-Draft-v1-F16.gguf filter=lfs diff=lfs merge=lfs -text
37
  LFM2.5-1.2B-Instruct-DSpark-Draft-v1-Q8_0.gguf filter=lfs diff=lfs merge=lfs -text
38
  LFM2.5-1.2B-Instruct-DSpark-Draft-v1-Q4_K_M.gguf filter=lfs diff=lfs merge=lfs -text
39
+ LFM2.5-1.2B-Instruct-DSpark-F16.gguf filter=lfs diff=lfs merge=lfs -text
40
+ LFM2.5-1.2B-Instruct-DSpark-Q8_0.gguf filter=lfs diff=lfs merge=lfs -text
41
+ LFM2.5-1.2B-Instruct-DSpark-Q4_K_M.gguf filter=lfs diff=lfs merge=lfs -text
LFM2.5-1.2B-Instruct-DSpark-Draft-v1-F16.gguf β†’ LFM2.5-1.2B-Instruct-DSpark-F16.gguf RENAMED
File without changes
LFM2.5-1.2B-Instruct-DSpark-Draft-v1-Q4_K_M.gguf β†’ LFM2.5-1.2B-Instruct-DSpark-Q4_K_M.gguf RENAMED
File without changes
LFM2.5-1.2B-Instruct-DSpark-Draft-v1-Q8_0.gguf β†’ LFM2.5-1.2B-Instruct-DSpark-Q8_0.gguf RENAMED
File without changes
README.md CHANGED
@@ -1,6 +1,6 @@
1
  ---
2
  library_name: llama.cpp
3
- base_model: LiquidAI/LFM2.5-1.2B-Instruct-DSpark-Draft-v1
4
  license: other
5
  license_name: lfm1.0
6
  license_link: LICENSE
@@ -39,9 +39,9 @@ Find more information about LFM2.5-DSpark in our [blog post](https://www.liquid.
39
 
40
  | file | quant | size | notes |
41
  |---|---|---:|---|
42
- | `LFM2.5-1.2B-Instruct-DSpark-Draft-v1-F16.gguf` | F16 | 594 MB | best accept length, recommended when memory allows |
43
- | `LFM2.5-1.2B-Instruct-DSpark-Draft-v1-Q8_0.gguf` | Q8_0 | 315 MB | accept length βˆ’2% vs F16 |
44
- | `LFM2.5-1.2B-Instruct-DSpark-Draft-v1-Q4_K_M.gguf` | Q4_K_M | 174 MB | accept length βˆ’3% vs F16, smallest recommended β€” sub-4-bit draft quants measurably hurt both accept length and throughput |
45
 
46
  Draft quantization changes speed only marginally (the drafter is a small share of each cycle); choose by memory budget. The target model quant is the main speed/quality lever and is independent of this file.
47
 
@@ -49,7 +49,7 @@ Draft quantization changes speed only marginally (the drafter is a small share o
49
 
50
  ```bash
51
  llama-server -m LFM2.5-1.2B-Instruct-F16.gguf \
52
- -md LFM2.5-1.2B-Instruct-DSpark-Draft-v1-F16.gguf \
53
  --spec-type draft-dspark --spec-draft-n-max 10 --spec-draft-n-min 0 \
54
  -fa on -ngl 99
55
  ```
 
1
  ---
2
  library_name: llama.cpp
3
+ base_model: LiquidAI/LFM2.5-1.2B-Instruct-DSpark
4
  license: other
5
  license_name: lfm1.0
6
  license_link: LICENSE
 
39
 
40
  | file | quant | size | notes |
41
  |---|---|---:|---|
42
+ | `LFM2.5-1.2B-Instruct-DSpark-F16.gguf` | F16 | 594 MB | best accept length, recommended when memory allows |
43
+ | `LFM2.5-1.2B-Instruct-DSpark-Q8_0.gguf` | Q8_0 | 315 MB | accept length βˆ’2% vs F16 |
44
+ | `LFM2.5-1.2B-Instruct-DSpark-Q4_K_M.gguf` | Q4_K_M | 174 MB | accept length βˆ’3% vs F16, smallest recommended β€” sub-4-bit draft quants measurably hurt both accept length and throughput |
45
 
46
  Draft quantization changes speed only marginally (the drafter is a small share of each cycle); choose by memory budget. The target model quant is the main speed/quality lever and is independent of this file.
47
 
 
49
 
50
  ```bash
51
  llama-server -m LFM2.5-1.2B-Instruct-F16.gguf \
52
+ -md LFM2.5-1.2B-Instruct-DSpark-F16.gguf \
53
  --spec-type draft-dspark --spec-draft-n-max 10 --spec-draft-n-min 0 \
54
  -fa on -ngl 99
55
  ```