Plans to upstream this patch in vLLM?

#2
by ibaldonl - opened

Are there any plans to upstream this patch into vLLM itself?

I searched issues and pull requests for "GSQ" which didn't found anything.

IST Austria Distributed Algorithms and Systems Lab org

Thanks for bringing this to our attention. We don’t know why vllm has not added the support for embedding quantization. But I will open an issue there.

Sign up or log in to comment