Mã QR

DGPE: A Dual-Stream Geometric-Aware Positional Encoding for Transformer-Based Object Detection

Accurately modeling complex spatial relationships is a pivotal challenge for Transformer-based object detectors, as existing models often rely on static positional encodings that fail to capture dynamic, real-world scene layouts. To address this issue, this work proposes a novel Dual-stream Geometri...

Mô tả đầy đủ

Đã lưu trong:
Chi tiết về thư mục
Những tác giả chính: Han-Cheng Hsiang, Zunliang Yan, Yong Liu
Định dạng: Artigo
Ngôn ngữ:Inglês
Được phát hành: IEEE 2025-01-01
Loạt:IEEE Access
Những chủ đề:
Truy cập trực tuyến:https://ieeexplore.ieee.org/document/11278171/
Các nhãn: Thêm thẻ
Không có thẻ, Là người đầu tiên thẻ bản ghi này!