Bunch of bad models
Sawyer Bowerman
soyrsoyr
AI & ML interests
None yet
Recent Activity
updated a model about 11 hours ago
soyrsoyr/Qwen3.6-27B-FP8-dynamic published a model about 11 hours ago
soyrsoyr/Qwen3.6-27B-FP8-dynamic new activity 3 days ago
huggingface/InferenceSupport:RedHatAI/DeepSeek-R1-Distill-Qwen-32B-quantized.w4a16Organizations
Llama-3.2-1B-Instruct GPTQ Quantized
GPTQ quantized across W4A16, W8A8, FP8, NVFP4 using llm-compressor.
-
soyrsoyr/Llama-3.2-1B-Instruct-W4A16-GPTQ
Text Generation • 1B • Updated • 6 -
soyrsoyr/Llama-3.2-1B-Instruct-W8A8-GPTQ
Text Generation • 1B • Updated • 14 -
soyrsoyr/Llama-3.2-1B-Instruct-FP8-GPTQ
Text Generation • 1B • Updated • 8 -
soyrsoyr/Llama-3.2-1B-Instruct-NVFP4-GPTQ
Text Generation • 0.8B • Updated • 9
Gemma 4 12B Quantized
Quantized Models
Bunch of good models
DeepSeek-MoE-16B-Chat GPTQ Quantized
DeepSeek-MoE-16B-Chat quantized with GPTQ via llm-compressor: W8A8, W4A16, FP8, NVFP4.
-
soyrsoyr/deepseek-moe-16b-chat-W8A8-GPTQ
Text Generation • 16B • Updated • 11 -
soyrsoyr/deepseek-moe-16b-chat-W4A16-GPTQ
Text Generation • 3B • Updated • 10 -
soyrsoyr/deepseek-moe-16b-chat-FP8-GPTQ
Text Generation • 16B • Updated • 4 -
soyrsoyr/deepseek-moe-16b-chat-NVFP4-GPTQ
Text Generation • 9B • Updated • 9
Custom Models
Bunch of bad models
Quantized Models
Bunch of good models
Llama-3.2-1B-Instruct GPTQ Quantized
GPTQ quantized across W4A16, W8A8, FP8, NVFP4 using llm-compressor.
-
soyrsoyr/Llama-3.2-1B-Instruct-W4A16-GPTQ
Text Generation • 1B • Updated • 6 -
soyrsoyr/Llama-3.2-1B-Instruct-W8A8-GPTQ
Text Generation • 1B • Updated • 14 -
soyrsoyr/Llama-3.2-1B-Instruct-FP8-GPTQ
Text Generation • 1B • Updated • 8 -
soyrsoyr/Llama-3.2-1B-Instruct-NVFP4-GPTQ
Text Generation • 0.8B • Updated • 9
DeepSeek-MoE-16B-Chat GPTQ Quantized
DeepSeek-MoE-16B-Chat quantized with GPTQ via llm-compressor: W8A8, W4A16, FP8, NVFP4.
-
soyrsoyr/deepseek-moe-16b-chat-W8A8-GPTQ
Text Generation • 16B • Updated • 11 -
soyrsoyr/deepseek-moe-16b-chat-W4A16-GPTQ
Text Generation • 3B • Updated • 10 -
soyrsoyr/deepseek-moe-16b-chat-FP8-GPTQ
Text Generation • 16B • Updated • 4 -
soyrsoyr/deepseek-moe-16b-chat-NVFP4-GPTQ
Text Generation • 9B • Updated • 9
Gemma 4 12B Quantized