Scale-QLoRA: Code-invariant adapter merging for native 4-bit microscaling LLMs
Read the original at arxiv.org→arXiv:2609.04526v1 Announce Type: new Abstract: Merging a LoRA adapter into its base model is standard deployment practice: it removes the runtime adapter's per-forward overhead and leaves a single standalone...
Original headline: "Scale-QLoRA: Code-Invariant Adapter Merging for Native 4-bit Microscaling LLMs"
Coverage timeline
- Sep 7, 04:00 UTC arXiv cs.CL lead source Scale-QLoRA: Code-Invariant Adapter Merging for Native 4-bit Microscaling LLMs