DGPE: A Dual-Stream Geometric-Aware Positional Encoding for Transformer-Based Object Detection
Accurately modeling complex spatial relationships is a pivotal challenge for Transformer-based object detectors, as existing models often rely on static positional encodings that fail to capture dynamic, real-world scene layouts. To address this issue, this work proposes a novel Dual-stream Geometri...
Đã lưu trong:
| Những tác giả chính: | , , |
|---|---|
| Định dạng: | Artigo |
| Ngôn ngữ: | Inglês |
| Được phát hành: |
IEEE
2025-01-01
|
| Loạt: | IEEE Access |
| Những chủ đề: | |
| Truy cập trực tuyến: | https://ieeexplore.ieee.org/document/11278171/ |
| Các nhãn: |
Không có thẻ, Là người đầu tiên thẻ bản ghi này!
|
