Vision Transformer with Deformable Attention

Transformers have recently shown superior performances on various vision tasks. The large, sometimes even global, receptive field endows Transformer models with higher representation power over their CNN counterparts. Nevertheless, simply enlarging receptive field also gives rise to several concerns...

Celý popis

Uloženo v:

Podrobná bibliografie
Vydáno v:	Proceedings (IEEE Computer Society Conference on Computer Vision and Pattern Recognition. Online) s. 4784 - 4793
Hlavní autoři:	Xia, Zhuofan, Pan, Xuran, Song, Shiji, Li, Li Erran, Huang, Gao
Médium:	Konferenční příspěvek
Jazyk:	angličtina
Vydáno:	IEEE 01.06.2022
Témata:	Adaptation models categorization Computational modeling Computer vision Data models Deformable models grouping and shape analysis Predictive models Recognition: detection retrieval; Deep learning architectures and techniques; Representation learning; Segmentation Transformers
ISSN:	1063-6919
On-line přístup:	Získat plný text
Tagy:	Přidat tag Žádné tagy, Buďte první, kdo vytvoří štítek k tomuto záznamu!

Buďte první, kdo okomentuje tento záznam!