bnjmnmarie commited on
Commit
1f66e17
·
verified ·
1 Parent(s): 55ef8d8

Update README.md

Browse files
Files changed (1) hide show
  1. README.md +5 -0
README.md CHANGED
@@ -8,6 +8,11 @@ tags:
8
  This is [allenai/Olmo-3-32B-Think](https://huggingface.co/allenai/Olmo-3-32B-Think) quantized with [LLM Compressor](https://github.com/vllm-project/llm-compressor) with NVFP4. The model has been created, tested, and evaluated by The Kaitchup.
9
  The model is compatible with vLLM (tested: v0.11.1). Tested with an RTX 5090.
10
 
 
 
 
 
 
11
 
12
  - **Developed by:** [The Kaitchup](https://kaitchup.substack.com/)
13
  - **License:** Apache 2.0 license
 
8
  This is [allenai/Olmo-3-32B-Think](https://huggingface.co/allenai/Olmo-3-32B-Think) quantized with [LLM Compressor](https://github.com/vllm-project/llm-compressor) with NVFP4. The model has been created, tested, and evaluated by The Kaitchup.
9
  The model is compatible with vLLM (tested: v0.11.1). Tested with an RTX 5090.
10
 
11
+ How the models perform (token efficiency, accuracy per domain, ...) and how to use them:
12
+ [Quantizing Olmo 3: Most Efficient and Accurate Formats](https://kaitchup.substack.com/p/quantizing-olmo-3-most-efficient)
13
+
14
+ ![image](https://cdn-uploads.huggingface.co/production/uploads/64b93e6bd6c468ac7536607e/H3JWV_ha07IrN-Sz6C7VL.png)
15
+
16
 
17
  - **Developed by:** [The Kaitchup](https://kaitchup.substack.com/)
18
  - **License:** Apache 2.0 license