LGViT: A Local and Global Vision Transformer with Dynamic Contextual Position Bias Using Overlapping Windows
Vision Transformers (ViTs) have shown their superiority in various visual tasks for the capability of self-attention mechanisms to model long-range dependencies. Some recent works try to reduce the high cost of vision transformers by limiting the self-attention module in a local window. As a price,...
Na minha lista:
| Principais autores: | , , |
|---|---|
| Formato: | Artigo |
| Idioma: | Inglês |
| Publicado em: |
MDPI AG
2023-02-01
|
| coleção: | Applied Sciences |
| Assuntos: | |
| Acesso em linha: | https://www.mdpi.com/2076-3417/13/3/1993 |
| Tags: |
Sem tags, seja o primeiro a adicionar uma tag!
|
