MetaTOC stay on top of your field, easily

Sequence length scaling in vision transformers for scientific images on frontier

, , , , , , , , , , , , , , , , , , , , , , ,

The International Journal of High Performance Computing Applications

Published online on

Abstract

The International Journal of High Performance Computing Applications, Volume 40, Issue 3, Page 273-290, May 2026.
Vision Transformers (ViTs) are pivotal for foundational models in scientific imagery, including Earth science applications, due to their capability to process large sequence lengths. While transformers for text have inspired scaling sequence lengths in ...