File size: 3,687 Bytes
a63f8ce
 
01dbe7b
 
 
 
 
 
 
 
 
 
 
 
 
 
 
a63f8ce
 
 
 
01dbe7b
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
a63f8ce
dafb389
 
 
 
 
 
 
8cf4787
dafb389
 
 
 
 
 
8cf4787
 
 
 
 
 
 
 
 
 
 
 
 
dafb389
a63f8ce
519040b
a63f8ce
519040b
a63f8ce
519040b
8cf4787
519040b
a63f8ce
519040b
a63f8ce
519040b
8cf4787
519040b
8cf4787
519040b
8cf4787
519040b
8cf4787
 
 
519040b
 
 
 
8cf4787
a63f8ce
 
519040b
a63f8ce
519040b
a63f8ce
 
519040b
 
 
a63f8ce
519040b
a63f8ce
519040b
 
 
 
 
8cf4787
 
a63f8ce
519040b
1
2
3
4
5
6
7
8
9
10
11
12
13
14
15
16
17
18
19
20
21
22
23
24
25
26
27
28
29
30
31
32
33
34
35
36
37
38
39
40
41
42
43
44
45
46
47
48
49
50
51
52
53
54
55
56
57
58
59
60
61
62
63
64
65
66
67
68
69
70
71
72
73
74
75
76
77
78
79
80
81
82
83
84
85
86
87
88
89
90
91
92
93
94
95
96
97
98
99
100
101
102
103
104
105
106
107
108
109
110
111
112
113
114
115
116
117
118
119
120
121
122
123
124
125
126
127
128
129
130
131
132
133
134
135
136
137
138
139
140
141
142
---
language:
- en
- zh
- fr
- es
- pt
- de
- it
- ru
- ja
- ko
- ar
- vi
- th
- nl
- pl
license: apache-2.0
library_name: mlx
base_model: Qwen/Qwen3.5-9B
tags:
- 16gb-mac
- 4-bit
- 4bit
- apple-silicon
- balanced
- chat
- conversational
- edge-ai
- everyday
- function-calling
- instruct
- lite
- local-llm
- m1
- m2
- m3
- m4
- mac
- mac-mini
- mac-studio
- macbook-air
- macbook-pro
- macos
- metal
- mlx
- mlx-lm
- mmlu-verified
- no-cloud
- offline
- on-device
- outlier
- outlier-app
- private
- private-ai
- quantized
- qwen
- qwen3.5
- qwen3_5
- reasoning
- safetensors
- text-generation
- thinking
- tool-use
pipeline_tag: text-generation
model-index:
- name: Outlier-Ai/Outlier-Lite-9B-MLX-4bit
  results:
  - task:
      type: text-generation
      name: Text Generation
    dataset:
      name: MMLU (5-shot, n=14042)
      type: cais/mmlu
      config: all
      split: test
    metrics:
    - type: acc
      name: accuracy
      value: 0.7846
      verified: false
  - task:
      type: text-generation
      name: Text Generation
    dataset:
      name: HumanEval
      type: openai_humaneval
      split: test
    metrics:
    - type: pass@1
      name: pass@1
      value: 0.6951
      verified: false
---
> **Part of the [Outlier](https://outlier.host/?utm_source=hf&utm_medium=modelcard&utm_campaign=outlier_lite_9b_mlx_4bit) shipping lineup.** Outlier is a free macOS app that runs this model locally, with one click. Apple Silicon only.

# Outlier Lite 9B (MLX 4-bit)

The mid-tier shipping model. 9B dense, text-only, MLX 4-bit. Recommended for 12 GB+ Macs that want stronger reasoning than Nano without paying the latency cost of a 27B-class model.

## Try it in Outlier

The simplest way to use this model is through the Outlier app — open the tier picker, select **Outlier Lite**, click download, and chat. No setup, no Python, no MLX install, no token quotas.

➡ **[Download Outlier — outlier.host](https://outlier.host/?utm_source=hf&utm_medium=modelcard&utm_campaign=outlier_lite_9b_mlx_4bit)**

A screenshot of the tier picker is at [outlier.host/screenshots/tier-picker.png](https://outlier.host/screenshots/tier-picker.png?utm_source=hf&utm_medium=modelcard&utm_campaign=outlier_lite_9b_mlx_4bit).

## Load this directly (power users)

If you want the raw MLX-4bit weights without the app:

```bash
pip install mlx-lm
python -m mlx_lm.generate \
  --model Outlier-Ai/Outlier-Lite-9B-MLX-4bit \
  --prompt "Write a quicksort in Python." \
  --max-tokens 512
```

```python
from mlx_lm import load, generate
model, tokenizer = load("Outlier-Ai/Outlier-Lite-9B-MLX-4bit")
print(generate(model, tokenizer, prompt="Hello", max_tokens=256))
```

## Verified benchmarks

For σ-qualified MMLU, HumanEval, and Mac inference-speed numbers — with full provenance (source file, command, n, stderr, date) — see **[outlier.host/benchmarks](https://outlier.host/benchmarks?utm_source=hf&utm_medium=modelcard&utm_campaign=outlier_lite_9b_mlx_4bit)**.

## Other Outlier shipping tiers

- [Outlier Nano 4B (entry tier, ~3 GB)](https://huggingface.co/Outlier-Ai/Outlier-Nano-4B-MLX-4bit)
- [Outlier Quick 26B-A4B MoE (~16 GB)](https://huggingface.co/Outlier-Ai/Outlier-Quick-26B-MLX-4bit)
- [Outlier Core 27B (default, ~16 GB)](https://huggingface.co/Outlier-Ai/Outlier-Core-27B-MLX-4bit)
- [Outlier Code 27B (code-tuned, ~16 GB)](https://huggingface.co/Outlier-Ai/Outlier-Code-27B-MLX-4bit)
- [Outlier Vision 35B-A3B (multimodal, ~20 GB)](https://huggingface.co/Outlier-Ai/Outlier-Vision-35B-A3B-MLX-4bit)

## License

Apache 2.0 (inherits from upstream base model). Conversion artifact only — the underlying weights are governed by the base model's license.