Vision Transformer with Deformable Attention

Transformers have recently shown superior performances on various vision tasks. The large, sometimes even global, receptive field endows Transformer models with higher representation power over their CNN counterparts. Nevertheless, simply enlarging receptive field also gives rise to several concerns...

Full description

Saved in:

Bibliographic Details
Published in:	Proceedings (IEEE Computer Society Conference on Computer Vision and Pattern Recognition. Online) pp. 4784 - 4793
Main Authors:	Xia, Zhuofan, Pan, Xuran, Song, Shiji, Li, Li Erran, Huang, Gao
Format:	Conference Proceeding
Language:	English
Published:	IEEE 01.06.2022
Subjects:	Adaptation models categorization Computational modeling Computer vision Data models Deformable models grouping and shape analysis Predictive models Recognition: detection retrieval; Deep learning architectures and techniques; Representation learning; Segmentation Transformers
ISSN:	1063-6919
Online Access:	Get full text
Tags:	Add Tag No Tags, Be the first to tag this record!

Be the first to leave a comment!