Cap4Bridge: Caption-Guided Cross-Modal Contextualization With Stochastic Augmentation for Text-Video Retrieval
A key challenge in text-video retrieval is bridging the semantic gap between information-rich videos and concise text queries. Existing methods often address this by incorporating auxiliary captions from Large Language Models (LLMs) or employing stochastic modeling. However, these approaches face cr...
சேமிக்கப்பட்டது:
| முதன்மை ஆசிரியர்கள்: | , , , , , |
|---|---|
| வடிவம்: | Artigo |
| மொழி: | Inglês |
| வெளியிடப்பட்டது: |
IEEE
2026-01-01
|
| தொடர்: | IEEE Access |
| பொருள்கள்: | |
| ஆன்லைன் அணுகல்: | https://ieeexplore.ieee.org/document/11474843/ |
| குறிச்சொற்கள்: |
டாக்ஸ் இல்லை, இந்த பதிவுக்கு குறிச்சொல் சேர்க்கும் முதல் நபராக இருங்கள்!
|
