Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
88plug 's Collections
88plug All Quantizations
88plug W4A16 Quantizations
88plug W8A16 Quantizations

88plug W4A16 Quantizations

updated Jul 16

INT4 weight-only (W4A16) compressed-tensors. Gold path: AutoRound iters=200 where stamped. Native vLLM.

Upvote
-

  • 88plug/Qwen3.6-35B-A3B-W4A16

    Image-Text-to-Text • 35B • Updated Jul 14 • 1.95k

  • 88plug/Qwen3.6-27B-W4A16

    Image-Text-to-Text • 27B • Updated Jul 15 • 1.21k

  • 88plug/Qwen3-Omni-30B-W4A16

    Text Generation • 35B • Updated Jul 10 • 156 • 1

  • 88plug/Qwen2.5-Omni-7B-W4A16

    Text Generation • 11B • Updated Jul 17 • 82

  • 88plug/Gemma4-E4B-it-W4A16

    Image-Text-to-Text • 8B • Updated Jul 11 • 63

  • 88plug/Gemma4-E2B-it-W4A16

    Image-Text-to-Text • 5B • Updated Jul 20 • 349

  • 88plug/MiniCPM-o-4.5-W4A16

    Image-Text-to-Text • 9B • Updated Jul 14 • 49 • 2

  • 88plug/Nemotron-3-Nano-30B-A3B-W4A16

    Text Generation • 33B • Updated Jul 14 • 124

  • 88plug/Kimi-VL-A3B-Thinking-2506-W4A16

    Image-Text-to-Text • 17B • Updated Jul 10 • 110
Upvote
-
  • Collection guide
  • Browse collections
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs