DreamFoundries commited on
Commit
86faa47
·
verified ·
1 Parent(s): 9bc491f

Add files using upload-large-folder tool

Browse files
Files changed (3) hide show
  1. README.md +4 -2
  2. model.safetensors +2 -2
  3. model.safetensors.index.json +1 -3
README.md CHANGED
@@ -24,11 +24,13 @@ This repository contains an MLX-LM conversion of [google/gemma-4-E2B-it](https:/
24
  - Quantization: MLX-LM affine quantization
25
  - Bits: 6-bit
26
  - Group size: 64
27
- - Local MLX folder size at upload time: 3.53 GiB
28
- - Local safetensors weight size at upload time: 3.50 GiB
29
 
30
  This Gemma conversion follows the MLX-LM Gemma 4 shared-KV topology and uses non-strict checkpoint loading so extra HF tensors outside that topology are discarded during conversion.
31
 
 
 
32
  ## Usage
33
 
34
  ```bash
 
24
  - Quantization: MLX-LM affine quantization
25
  - Bits: 6-bit
26
  - Group size: 64
27
+ - Local MLX folder size at upload time: 3.55 GiB
28
+ - Local safetensors weight size at upload time: 3.52 GiB
29
 
30
  This Gemma conversion follows the MLX-LM Gemma 4 shared-KV topology and uses non-strict checkpoint loading so extra HF tensors outside that topology are discarded during conversion.
31
 
32
+ For mlx-swift compatibility, `per_layer_model_projection` was left unquantized while the rest of the eligible linear layers were quantized.
33
+
34
  ## Usage
35
 
36
  ```bash
model.safetensors CHANGED
@@ -1,3 +1,3 @@
1
  version https://git-lfs.github.com/spec/v1
2
- oid sha256:6d65ec9c8929c5f59c70696f3c95bc33845d2189614c563e3699a8e91c9eaa7f
3
- size 3761193772
 
1
  version https://git-lfs.github.com/spec/v1
2
+ oid sha256:10b198392de58d1dc0fedb3d0a62aeae323c13e2cd67cabb40f5e565a4ff2f89
3
+ size 3777536556
model.safetensors.index.json CHANGED
@@ -1,6 +1,6 @@
1
  {
2
  "metadata": {
3
- "total_size": 3761052230,
4
  "total_parameters": 4628569344
5
  },
6
  "weight_map": {
@@ -1096,8 +1096,6 @@
1096
  "language_model.model.layers.9.self_attn.v_proj.scales": "model.safetensors",
1097
  "language_model.model.layers.9.self_attn.v_proj.weight": "model.safetensors",
1098
  "language_model.model.norm.weight": "model.safetensors",
1099
- "language_model.model.per_layer_model_projection.biases": "model.safetensors",
1100
- "language_model.model.per_layer_model_projection.scales": "model.safetensors",
1101
  "language_model.model.per_layer_model_projection.weight": "model.safetensors",
1102
  "language_model.model.per_layer_projection_norm.weight": "model.safetensors"
1103
  }
 
1
  {
2
  "metadata": {
3
+ "total_size": 3777395270,
4
  "total_parameters": 4628569344
5
  },
6
  "weight_map": {
 
1096
  "language_model.model.layers.9.self_attn.v_proj.scales": "model.safetensors",
1097
  "language_model.model.layers.9.self_attn.v_proj.weight": "model.safetensors",
1098
  "language_model.model.norm.weight": "model.safetensors",
 
 
1099
  "language_model.model.per_layer_model_projection.weight": "model.safetensors",
1100
  "language_model.model.per_layer_projection_norm.weight": "model.safetensors"
1101
  }