Running on vLLM

#6
by socrat3k - opened

Hi,
I understand GGUF is poorly-supported on vLLM, so it's expected that I am unable to run the model natively there. I have then a question, is there a way I can reliably run it on vLLM?

Sign up or log in to comment