CVTrack: Combined Convolutional Neural Network and Vision Transformer Fusion Model for Visual Tracking
Most single-object trackers currently employ either a convolutional neural network (CNN) or a vision transformer as the backbone for object tracking. In CNNs, convolutional operations excel at extracting local features but struggle to capture global representations. On the other hand, vision transfo...
保存先:
| 主要な著者: | , , , , , |
|---|---|
| フォーマット: | Artigo |
| 言語: | Inglês |
| 出版事項: |
MDPI AG
2024-01-01
|
| シリーズ: | Sensors |
| 主題: | |
| オンライン・アクセス: | https://www.mdpi.com/1424-8220/24/1/274 |
| タグ: |
タグなし, このレコードへの初めてのタグを付けませんか!
|
