SPT-Swin: A Shifted Patch Tokenization Swin Transformer for Image Classification
Recently, the transformer-based model e.g., the vision transformer (ViT) has been extensively used in computer vision tasks. The superior performance of the ViT leads to the requirement of an enormous dataset and the complexity of calculating self-attention between patches is quadratic in nature. To...
Gardado en:
| Principais autores: | , , , |
|---|---|
| Formato: | Artigo |
| Idioma: | Inglês |
| Publicado: |
IEEE
2024-01-01
|
| Series: | IEEE Access |
| Assuntos: | |
| Acceso en liña: | https://ieeexplore.ieee.org/document/10643534/ |
| Tags: |
Sen Etiquetas, Sexa o primeiro en etiquetar este rexistro!
|
