Preprint2023
ConvNets Match Vision Transformers at Scale
Unknown
ConvNets (NFNets) match Vision Transformers in performance when pre-trained on large datasets and fine-tuned on ImageNet, challenging the assumption that ViTs are inherently superior at scale.
0Oct 1, 2023TransformersFine Tuning