BriefGPT.xyz
Apr, 2021
Visformer:基于Transformer的视觉友好型模型
Visformer: The Vision-friendly Transformer
HTML
PDF
Zhengsu Chen, Lingxi Xie, Jianwei Niu, Xuefeng Liu, Longhui Wei...
TL;DR
通过将Transformer模型逐步转化为基于卷积的模型,进行经验证实验表明,命名为Visformer的新架构比Transformer模型和卷积模型在ImageNet分类准确性方面表现更佳,尤其是当模型复杂度较低或者训练集较小时。
Abstract
The past year has witnessed the rapid development of applying the
transformer module
to
vision problems
. While some researchers have demonstrated that Transformer-based models enjoy a favorable ability of fitting
→