Original model by: Qwen
Original model: Qwen3.5-2B

For more information about the model, check out the original model page and the creator while you're at it.

ExLlamaV3 Quantizations:
8.0bpw: 8hb
6.0bpw: 6hb
5.0bpw: 6hb
4.0bpw: 6hb
3.0bpw: 6hb
2.5bpw: 6hb
2.25bpw: 6hb
2.0bpw: 6hb

Made with EXL3 Version: 1.4.1

If you need a specific bits-per-weight or head-bits combo, please let me know. I’m happy to help.

Your feedback and suggestions are always welcome! They help me improve and make quantizations better for everyone.

Special thanks to turboderp for developing the tools that made these quantizations possible. Your contributions are greatly appreciated!

Downloads last month

-

Downloads are not tracked for this model. How to track
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for TheMelonGod/Qwen3.5-2B-exl3

Finetuned
Qwen/Qwen3.5-2B
Quantized
(175)
this model