Código QR

SPT-Swin: A Shifted Patch Tokenization Swin Transformer for Image Classification

Recently, the transformer-based model e.g., the vision transformer (ViT) has been extensively used in computer vision tasks. The superior performance of the ViT leads to the requirement of an enormous dataset and the complexity of calculating self-attention between patches is quadratic in nature. To...

Descrición completa

Gardado en:
Detalles Bibliográficos
Principais autores: Gazi Jannatul Ferdous, Khaleda Akhter Sathi, Md. Azad Hossain, M. Ali Akber Dewan
Formato: Artigo
Idioma:Inglês
Publicado: IEEE 2024-01-01
Series:IEEE Access
Assuntos:
Acceso en liña:https://ieeexplore.ieee.org/document/10643534/
Tags: Engadir etiqueta
Sen Etiquetas, Sexa o primeiro en etiquetar este rexistro!