LS-LoRA changes only some layers and keeps more commonsense reasoning
It puts adapters only in layers with low input-output cosine similarity, and it increases the average target-task performance.
Claimed, not confirmed
This is a brief. We point to the report and do not rewrite it. Read it at the source below.
Sources
Posted