Aivora
arXivAI ResearchAdvanced

New LoRA Skills Should Read but Never Write: READ Solves Adapter Fusion Interference

新增 LoRA 技能唯讀不寫:READ 解決多配接器融合干擾

2 min read
New LoRA Skills Should Read but Never Write: READ Solves Adapter Fusion Interference
The 30-second version

Merging multiple independently trained LoRA adapters often degrades performance due to parameter interference. This paper identifies two hidden culprits: arbitrary factorization coordinates and bidirectional coupling that overwrites old skills. To solve this, the authors propose READ (Read-only Expansion of Adapter Deltas). READ transforms each adapter into a balanced canonical form and enforces a strict one-way coupling: new skills can only read old skills' input subspaces but never write to their outputs. READ outperforms strong baselines by over 20 points on SuperGLUE and 7 points on domain benchmarks.

Key points

01

Identifying Interference Root Causes

The study reveals that interference in merged LoRAs stems from arbitrary factorization coordinates and bidirectional coupling that disrupts pre-existing skills.

02

Balanced Canonical Form

READ rewrites each adapter into a balanced canonical form, ensuring consistent coordinate baselines for interactions between different adapters.

03

One-Way Read-Only Coupling

New skills can only read the input subspaces of old skills but are strictly blocked from writing to their output subspaces, avoiding degradation.

04

Zero Inference Overhead

The composed updates can be folded directly back into the base model weights, eliminating the need for runtime routing or extra computation.

How it works

Comparison between Traditional LoRA Merging and READ
傳統 LoRA 合併 (Traditional)READ (本研究方法)
Coupling Direction雙向耦合(容易互相覆寫與干擾)單向唯讀(新技能唯讀舊技能,不干擾舊輸出)
Factorization任意、不一致的座標系統標準化平衡型式 (Balanced Canonical Form)
Inference Cost高(若使用路由/MoE)或 表現不佳(直接相加)零成本(可直接摺疊融入基礎模型權重中)
Skill Retention差(技能易在合併過程中流失)極佳(舊技能計算路徑完全不受影響)

Why it matters

Traditionally, combining multiple specialized skills in an LLM required expensive multi-task retraining or complex MoE-style routing that added inference latency. READ offers an elegant mathematical solution to "stack" or "hot-plug" new skills at near-zero cost without damaging existing capabilities. This is highly valuable for enterprise LLM deployments requiring continuous learning and modular feature expansion.

Who it affects

  • AI Developer
  • AI Researcher
  • Enterprise Leader

How to use it

  1. 1Modular AI Skill Stacking
  2. 2Lifelong Learning without Catastrophic Forgetting

Limitations & caveats

  • Requires access to and transformation of the original LoRA adapter weights
  • Primarily evaluated on sequential skill addition; behavior with extremely high numbers of concurrent adapters is untested

Related