Any recipe for vLLM?

#1
by hdnh2006 - opened

Hello!

Thanks for this fantastic release, any idea how to deploy it in 24GB of VRAM using vLLM omni? is this even possible?

Thanks in advance.

ModelsLab org

vLLM omni won't support this format, use custom code stack, we have in readme, it will work. or better ask claude to make it work, it will work.

Sign up or log in to comment