QRコード

CVTrack: Combined Convolutional Neural Network and Vision Transformer Fusion Model for Visual Tracking

Most single-object trackers currently employ either a convolutional neural network (CNN) or a vision transformer as the backbone for object tracking. In CNNs, convolutional operations excel at extracting local features but struggle to capture global representations. On the other hand, vision transfo...

詳細記述

保存先:
書誌詳細
主要な著者: Jian Wang, Yueming Song, Ce Song, Haonan Tian, Shuai Zhang, Jinghui Sun
フォーマット: Artigo
言語:Inglês
出版事項: MDPI AG 2024-01-01
シリーズ:Sensors
主題:
オンライン・アクセス:https://www.mdpi.com/1424-8220/24/1/274
タグ: タグ追加
タグなし, このレコードへの初めてのタグを付けませんか!