OmDet: Large‐scale vision‐language multi‐dataset pre‐training with multimodal detection network
Abstract The advancement of object detection (OD) in open‐vocabulary and open‐world scenarios is a critical challenge in computer vision. OmDet, a novel language‐aware object detection architecture and an innovative training mechanism that harnesses continual learning and multi‐dataset vision‐langua...
محفوظ في:
| المؤلفون الرئيسيون: | , , |
|---|---|
| التنسيق: | Artigo |
| اللغة: | Inglês |
| منشور في: |
Wiley
2024-08-01
|
| سلاسل: | IET Computer Vision |
| الموضوعات: | |
| الوصول للمادة أونلاين: | https://doi.org/10.1049/cvi2.12268 |
| الوسوم: |
لا توجد وسوم, كن أول من يضع وسما على هذه التسجيلة!
|
