Live page ยท Day archive

V-CoLA: Vision Token Compression with Linear Attention

Read the original at HF Daily Papers

Summary

Researchers propose V-CoLA, a training-free framework that compresses vision tokens for linear attention models, achieving 99.5% of baseline performance.

Carried by: HF Daily Papers. First seen: .