README.md

January 28, 2025 ยท View on GitHub

Model Weights and Configurations

ModelToken GridTop-1 Acc.Config
FastChannelVim-S/16.ckpt142 x 873.6FastChannelVim-S/16.yaml
FastChannelVim-S/16 - Maxpool.ckpt142 x 872.9FastChannelVim-S/16 - Maxpool.yaml
ChannelVim-S/16.ckpt142 x 873.5ChannelVim-S/16.yaml
FastChannelVim-S/8.ckpt282 x 883.1FastChannelVim-S/8.yaml
FastChannelVim-S/8 - Maxpool.ckpt282 x 885.0FastChannelVim-S/8 - Maxpool.yaml
ChannelVim-S/8.ckpt282 x 883.0ChannelVim-S/8.yaml

Notes:

  • For reproducibility, make sure overall batch size remains 256 across GPUs/Nodes.
  • The preprocessed JUMP-CP data used in this paper was previously released along with "Contextual Vision Transformers for Robust Representation Learning" insitro/ContextViT.