Artificial Intelligence #finetuning#vision-language-action
New Training-Free Method Compresses Vision-Language-Action Models by 50% Without Performance Loss
A research team led by Gia-Binh Ho et al. discovered that Vision-Language-Action (VLA) models exhibit severe layer-wise redundancy. They introduced a training-free compression pipeline using Centered Kernel Alignment to remove twin layers, achieving up to 50% depth reduction, 40-50% faster fine-tuning, and 30% faster inference while matching or exceeding full-scale performance.
Jun 20, 2026 1 source