Huihui ThinkingCap Qwen3.6 27B Abliterated — MLX 4-bit MTP

MTPLX-compatible conversion of huihui-ai/Huihui-ThinkingCap-Qwen3.6-27B-abliterated, pinned to revision 44f63da.

Runtime-specific artifact: do not load this repository in LM Studio. LM Studio flattens the MTPLX sidecar into its target directory, causing the target loader to reject the 15 mtp.* tensors. For the recommended MTP experience, use the oMLX Native-MTP model; oMLX is the faster, more mature integrated path for this model.

  • Trunk: MLX affine 4-bit, group size 64.
  • MTP head: 15-tensor BF16 sidecar at mtp/weights.safetensors.
  • Runtime: MTPLX 2.1.0, Apple Silicon only.
pip install -U mtplx
mtplx run --model pixelkaiser/Huihui-ThinkingCap-Qwen3.6-27B-abliterated-MLX-4bit-MTP --depth 3 "Hello"

MTPLX Forge classified this artifact as verified-native. A bounded 32-token M4 Max verification measured 24.94 tok/s AR and 46.02 tok/s at MTP depth 3 (1.85x); treat this as a packaging smoke, not a general benchmark.

This is an abliterated model with reduced safety behavior. Review the upstream model card before use.

Downloads last month
638
Safetensors
Model size
5B params
Tensor type
BF16
·
U32
·
MLX
Hardware compatibility
Log In to add your hardware

4-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for pixelkaiser/Huihui-ThinkingCap-Qwen3.6-27B-abliterated-MLX-4bit-MTP