AIModels.fyi

AIModels.fyi

Researchers discover explicit registers eliminate vision transformer attention spikes

When visualizing the inner workings of vision transformers (ViTs), researchers noticed weird spikes of attention on random background patches. Here's how they fixed them.

aimodels-fyi's avatar
aimodels-fyi
Oct 01, 2023
∙ Paid
Researchers discover explicit registers eliminate vision transformer attention spikes
The impact of registers - getting ViTs to focus correctly

Transformers have become the model architecture of choice for many vision tasks. Vision Transformers (ViTs) are especially popular. They apply the transformer directly to sequences of image patches. ViTs now match or exceed CNNs on benchmarks like image classification.

However, researchers from Met…

User's avatar

Continue reading this post for free, courtesy of aimodels-fyi.

Or purchase a paid subscription.
© 2026 AIModels.fyi · Privacy ∙ Terms ∙ Collection notice
Start your SubstackGet the app
Substack is the home for great culture