
Model merging usually means hand-tuning parameters or running slow evolutionary search. DAM learns per-column scaling coefficients by gradient optimisation instead, reportedly matching or beating DARE-TIES and Model Soups at lower cost.
Checking sign-in…
Loading comments…