Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
caiovicentino1
/
GLM-4.7-Flash-HLWQ-Q5
like
1
glm4_moe_lite
hlwq
Mixture of Experts
glm
mla
bit-packed
polarengine
arxiv:
2502.02617
arxiv:
2603.29078
License:
mit
Model card
Files
Files and versions
xet
Community
Copy to bucket
new
main
GLM-4.7-Flash-HLWQ-Q5
19 GB
Ctrl+K
Ctrl+K
1 contributor
History:
17 commits
caiovicentino1
Remove legacy polar_config.json
0e69eb5
verified
4 months ago
assets
Upload assets/chart_download.png with huggingface_hub
5 months ago
.gitattributes
Safe
1.57 kB
Upload tokenizer.json with huggingface_hub
5 months ago
README.md
Safe
3.77 kB
HLWQ rebrand: title, tags, notice, self-links
4 months ago
chat_template.jinja
Safe
3.12 kB
Upload chat_template.jinja with huggingface_hub
5 months ago
config.json
Safe
2.08 kB
fix: quant_method polar -> polarengine for vLLM compatibility
5 months ago
hlwq_config.json
401 Bytes
Add hlwq_config.json (rename from polar_config.json)
4 months ago
polar_state.safetensors
Safe
19 GB
xet
Upload polar_state.safetensors with huggingface_hub
5 months ago
tokenizer.json
Safe
20.2 MB
xet
Upload tokenizer.json with huggingface_hub
5 months ago
tokenizer_config.json
Safe
305 Bytes
Upload tokenizer_config.json with huggingface_hub
5 months ago