Riemannian--Lorentz Fusion of Vision Transformers and State-Space Models

Researchers propose Riemannian--Lorentz Parameter Fusion (RLPF), a method for merging pre-trained Vision Transformers and state-space models by aligning parameter groups by semantic role, enabling orders-of-magnitude savings versus retraining. Initial results show improved accuracy on CIFAR-10, Oxford-IIIT Pet, and ImageNet-1K.

RSS Score 0 9/18/2026, 4:00:00 AM Original Source
Save an API key to vote.