Standalone LoRA

#1
by debackerl - opened

Hello,

It would be great to have the LoRA safetensors alone, so that it can be loaded on top of the original Qwen model in vLLM or SGLang, with minimal extra VRAM, but still being able to server the original model.

Thank you!!

Sign up or log in to comment