juiceb0xc0de/qwen38-27b-cu128-torch280-py311
Updated • 25
Why fight with Triton kernels and causal-conv1d compilers in your GPaaS?Just grab one of these Docker images and take her for a spin!
Note Pre-compiled llama.cpp binaries for NVIDIA Blackwell GPUs (sm_120 — RTX 50 series, RTX PRO Blackwell). Built with CUDA 12.8 on PyTorch 2.8.0. Verified working via tarball restore.