Text Generation
Transformers
Safetensors
PyTorch
nemotron_h
nvidia
nemotron-3
latent-moe
mtp
conversational
Eval Results

SGLang or TensorRT deployment on H200s

#6
by naamanm - opened

Hi, has anyone deployed this on SGLang or TenrorRT LLM on 8 x H200s ? SGLang apparently does not support the GPUs and TensorRT did not have the checkpoint coversion code

Sign up or log in to comment