10Eros-Max-Beta2-Pruned (GGUF)

This repository contains heavily optimized GGUF quantizations of TenStrip/10Eros-Max (specifically the 10Eros_Max_h3_fl2va_beta2_pruned version).

These models are designed to run in ComfyUI natively via ComfyUI-GGUF, allowing users with consumer GPUs (12GB - 16GB VRAM) to run this massive 40GB+ MiniMax-H3 model locally.

πŸ“¦ Available Quantizations

File Size Description VRAM Target
10Eros-Max-Beta2-Pruned-Q3_K_M.gguf 8.9 GB Extreme compression. Expect some degradation in fine details. <12GB
10Eros-Max-Beta2-Pruned-Q4_K_M.gguf 11.6 GB ⭐️ Recommended. Perfect balance of quality and size. Preserves text and physics incredibly well. 12GB - 16GB
10Eros-Max-Beta2-Pruned-Q4_K_S.gguf 11.6 GB Slightly smaller alternative to Q4_K_M. 12GB - 16GB
10Eros-Max-Beta2-Pruned-Q5_K_M.gguf 14.1 GB High fidelity. Great for complex geometry and micro-details. 16GB - 24GB
10Eros-Max-Beta2-Pruned-Q5_K_S.gguf 14.1 GB Slightly smaller alternative to Q5_K_M. 16GB - 24GB
10Eros-Max-Beta2-Pruned-Q6_K.gguf 16.7 GB Near-lossless visual quality. 24GB+
10Eros-Max-Beta2-Pruned-Q8_0.gguf 21.6 GB Maximum quality, nearly indistinguishable from original BF16. 24GB+

πŸš€ How to Use in ComfyUI

  1. Install Custom Node: Install ComfyUI-GGUF (by city96) via the ComfyUI Manager.
  2. Download Model: Download your preferred .gguf file (Q4_K_M is recommended).
  3. Placement: Place the downloaded .gguf file inside your ComfyUI/models/unet directory.
  4. Workflow: Replace your standard Load Diffusion Model node with the Unet Loader (GGUF) node.

βš™οΈ Recommended Inference Settings

Note: This is the Standard / Pruned Beta 2 version, NOT the Turbo version.

  • Steps: 20 to 30
  • CFG Scale: ~3.5 to 4.5
  • Sampler/Scheduler: Standard MiniMax flow matching settings (e.g., euler / simple or flowmatch)

πŸ“œ Original Model Information & Credits

All credit for the underlying model architecture, grafting methodology, and training goes to TenStrip and the MiniMax team.

From the Original Creator (TenStrip):

"Due to a ton of confusion I've let Claude compile a full MD on the grafting, including code and methodology as it was mostly used to create code, document, and manage the graft project while I tested, architected, and mixed. I was keeping that to myself since maybe it'd actually be nice to have something to myself but after a handful of ignorant comments I'm open sourcing everything from it except the scripts themselves. You can hand the md back to claude or any competent agent and graft in the same way or have them explain it until you understand. No more dumbassery allowed now."

"This is an evolving project subject to future fixes in H3 training. Training is clearly problematic, and my branch will rely on Sulphur H3 tunes. But I didn't want to wait for all that so I took the data from older models, grafting it in where the model needs it, grafting it to attn layers at a low level that doesn't disturb H3's visual or audio output quality. That's essentialy it."

License

The standard H3 community license applies (minimax-h3-community-license-agreement). Because this release now carries transferred character from LTX 2.3, Wan 2.2, and Krea 2, the community licenses for those source models apply as well to the portions of character that came from each.

Downloads last month
4,290
GGUF
Model size
20B params
Architecture
wan
Hardware compatibility
Log In to add your hardware

3-bit

4-bit

5-bit

6-bit

8-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. πŸ™‹ Ask for provider support

Model tree for Abiray/10Eros-Max-fl2va-Beta2-GGUF

Quantized
(5)
this model