Text-Guided Visual Representation Optimization for Sensor-Acquired Video Temporal Grounding
Video temporal grounding (VTG) aims to localize a semantically relevant temporal segment within an untrimmed video based on a natural language query. The task continues to face challenges arising from cross-modal semantic misalignment, which is largely attributed to redundant visual content in senso...
Na minha lista:
| Principais autores: | , , , |
|---|---|
| Format: | Artigo |
| Sprog: | Inglês |
| Udgivet: |
MDPI AG
2025-07-01
|
| Serier: | Sensors |
| Fag: | |
| Online adgang: | https://www.mdpi.com/1424-8220/25/15/4704 |
| Tags: |
Ingen Tags, Vær først til at tagge denne postø!
|
