Código QR (código de barras bidimensional)

LGViT: A Local and Global Vision Transformer with Dynamic Contextual Position Bias Using Overlapping Windows

Vision Transformers (ViTs) have shown their superiority in various visual tasks for the capability of self-attention mechanisms to model long-range dependencies. Some recent works try to reduce the high cost of vision transformers by limiting the self-attention module in a local window. As a price,...

ver descrição completa

Na minha lista:
Detalhes bibliográficos
Principais autores: Qian Zhou, Hua Zou, Huanhuan Wu
Formato: Artigo
Idioma:Inglês
Publicado em: MDPI AG 2023-02-01
coleção:Applied Sciences
Assuntos:
Acesso em linha:https://www.mdpi.com/2076-3417/13/3/1993
Tags: Adicionar Tag
Sem tags, seja o primeiro a adicionar uma tag!